Papers
Topics
Authors
Recent
Search
2000 character limit reached

Acyclic Graphical Continuous Lyapunov Models

Updated 14 July 2026
  • Acyclic GCLMs are graphical models that impose a DAG structure on drift matrices in continuous-time linear stochastic systems, ensuring global identifiability.
  • They leverage the continuous Lyapunov equation to encode dependencies, supporting both Gaussian and non-Gaussian formulations with higher-order cumulant analysis.
  • Practical estimation methods, including lasso and moment-matching, address high-dimensional challenges while exploiting acyclicity for robust model selection.

Searching arXiv for papers on acyclic graphical continuous Lyapunov models and related identifiability results. Acyclic graphical continuous Lyapunov models (GCLMs) are a subclass of graphical continuous Lyapunov models in which the directed support of the drift matrix is a directed acyclic graph (DAG). They model each observation as a cross-sectional draw from the stationary law of a stable continuous-time linear stochastic system, so that dependence is encoded through a continuous Lyapunov equation rather than through a structural equation solved at a single time point. In the Gaussian formulation, the stationary distribution is determined by the drift matrix and the noise covariance; in the non-Gaussian formulation, higher-order cumulants of the stationary law satisfy generalized Lyapunov equations and can identify edge weights beyond covariance information alone (Varando et al., 2020, Recke et al., 17 Mar 2026).

1. Stochastic formulation and graph semantics

In the Gaussian setting, a GCLM starts from a multivariate Ornstein–Uhlenbeck process

dX(t)=M (X(t)−a) dt+D dW(t),\mathrm dX(t)=M\,(X(t)-a)\,\mathrm dt + D\,\mathrm dW(t),

with drift matrix M∈Rp×pM\in\mathbb R^{p\times p}, volatility matrix D∈Rp×pD\in\mathbb R^{p\times p}, and C=DD⊤C=D D^\top. If MM is Hurwitz, then X(t)X(t) has a unique stationary Gaussian law N(a,Σ)N(a,\Sigma), and Σ\Sigma is the unique positive-definite solution of

MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.

A graphical continuous Lyapunov model imposes a zero pattern on MM through a directed graph M∈Rp×pM\in\mathbb R^{p\times p}0, with an edge M∈Rp×pM\in\mathbb R^{p\times p}1 indicating that M∈Rp×pM\in\mathbb R^{p\times p}2 may be nonzero; self-loops M∈Rp×pM\in\mathbb R^{p\times p}3 are always included so that diagonal entries remain free (Dettling et al., 2022).

For fixed M∈Rp×pM\in\mathbb R^{p\times p}4, the model can be written as

M∈Rp×pM\in\mathbb R^{p\times p}5

where M∈Rp×pM\in\mathbb R^{p\times p}6 denotes the stable matrices with the prescribed sparsity pattern (Dettling et al., 2022). In the acyclic case, one may topologically order the nodes so that M∈Rp×pM\in\mathbb R^{p\times p}7, making the off-diagonal support triangular after reordering; stability is then typically enforced through negative diagonal entries (Varando et al., 2020).

The broader non-Gaussian extension replaces Brownian noise by a Lévy process. In the formulation studied by Recke and Hansen, one observes a cross-sectional sample M∈Rp×pM\in\mathbb R^{p\times p}8 from the stationary solution of

M∈Rp×pM\in\mathbb R^{p\times p}9

where D∈Rp×pD\in\mathbb R^{p\times p}0 is stable, D∈Rp×pD\in\mathbb R^{p\times p}1 whenever D∈Rp×pD\in\mathbb R^{p\times p}2, and D∈Rp×pD\in\mathbb R^{p\times p}3 is a D∈Rp×pD\in\mathbb R^{p\times p}4-dimensional Lévy process. Under mild integrability theorems, the stationary law exists uniquely and admits the representation

D∈Rp×pD\in\mathbb R^{p\times p}5

This places acyclic GCLMs inside a larger class of stationary linear Markov-process models whose observational content is entirely cross-sectional (Recke et al., 17 Mar 2026).

2. Lyapunov equations, trek expansions, and covariance geometry

The defining covariance relation is the continuous Lyapunov equation

D∈Rp×pD\in\mathbb R^{p\times p}6

or, after vectorization,

D∈Rp×pD\in\mathbb R^{p\times p}7

This vectorized form is central both for identifiability arguments and for estimation procedures, because it turns the covariance constraints into linear equations in D∈Rp×pD\in\mathbb R^{p\times p}8 once D∈Rp×pD\in\mathbb R^{p\times p}9 is regarded as fixed (Hansen, 2024).

Hansen’s trek rule gives a graphical expansion of C=DD⊤C=D D^\top0 in terms of the drift and volatility parameters. In the general mixed-graph setting, if C=DD⊤C=D D^\top1 has spectral radius C=DD⊤C=D D^\top2, then

C=DD⊤C=D D^\top3

where C=DD⊤C=D D^\top4 ranges over treks, C=DD⊤C=D D^\top5 and C=DD⊤C=D D^\top6 count the left and right directed parts, C=DD⊤C=D D^\top7, and C=DD⊤C=D D^\top8 is the trek weight (Hansen, 2024). In the acyclic case, after rescaling so that C=DD⊤C=D D^\top9, the expansion becomes a finite polynomial over treks in the acyclic base graph:

MM0

Thus, for acyclic GCLMs, each covariance entry is a polynomial in the off-diagonal drift parameters and the volatility entries rather than an infinite series (Hansen, 2024).

The trek perspective clarifies how acyclic Lyapunov models differ from linear additive-noise DAG models. The Lyapunov trek rule carries the extra combinatorial factors MM1, and even in DAG cases the induced covariance geometry is not the same as in algebraic structural-equation models. In particular, the star-graph formulas in Hansen’s analysis show that the implied MM2 need not satisfy the rank-one tetrad constraints of classical factor models (Hansen, 2024).

For special acyclic drifts, explicit recursions are available. In the path-graph case with MM3, MM4, and MM5, the covariances satisfy

MM6

For certain sign patterns and diagonal MM7, Hansen also derives the lower bound

MM8

and in simple path models the marginal variances strictly increase along the topological order (Hansen, 2024). This suggests that acyclicity constrains not only identifiability but also qualitative variance propagation.

3. Identifiability from covariance in the Gaussian acyclic case

In the Gaussian setting with fixed MM9, identifiability asks whether the map

X(t)X(t)0

is injective. The fiber at X(t)X(t)1 is

X(t)X(t)2

The model is globally identifiable if every fiber is a singleton, generically identifiable if singleton fibers fail only on an algebraic exception set, and non-identifiable if all fibers are infinite (Dettling et al., 2022).

Dettling, Homs, Améndola, Drton, and Hansen prove the basic characterization: for arbitrary directed graphs with self-loops and fixed X(t)X(t)3, the parametrization is globally identifiable if and only if the graph has no directed X(t)X(t)4-cycle. Since every DAG is simple in this sense, every acyclic GCLM is globally identifiable (Dettling et al., 2022). In the diagonal-noise case, the same equivalence holds as a corollary.

The proof relies on vectorized rank conditions. Keeping only the symmetric equations yields a X(t)X(t)5 matrix X(t)X(t)6 such that

X(t)X(t)7

If X(t)X(t)8 denotes the submatrix formed by the columns corresponding to the allowed edges, then

X(t)X(t)9

and

N(a,Σ)N(a,\Sigma)0

For DAGs, a topological ordering makes N(a,Σ)N(a,\Sigma)1 block-upper-triangular with positive determinants on the diagonal blocks, which yields full column rank for all N(a,Σ)N(a,\Sigma)2 in the model (Dettling et al., 2022).

This covariance-based result distinguishes acyclic GCLMs from non-simple graphs, where directed N(a,Σ)N(a,\Sigma)3-cycles create immediate non-uniqueness. It also places acyclic GCLMs closer to identifiable sparse covariance parametrizations than to generic cyclic Lyapunov models. A frequent misconception is that covariance alone is intrinsically insufficient to identify direction in continuous-time linear systems; within the Gaussian GCLM framework that statement is false for DAG supports with known N(a,Σ)N(a,\Sigma)4, because acyclicity already guarantees global identifiability (Dettling et al., 2022).

4. Higher-order cumulants and non-Gaussian acyclic models

The non-Gaussian theory extends the second-order Lyapunov equation to cumulant tensors of arbitrary order. If N(a,Σ)N(a,\Sigma)5 is the N(a,Σ)N(a,\Sigma)6-th cumulant tensor of the driving Lévy noise and N(a,Σ)N(a,\Sigma)7 is the N(a,Σ)N(a,\Sigma)8-th cumulant tensor of the stationary law, then N(a,Σ)N(a,\Sigma)9 satisfies

Σ\Sigma0

Equivalently, in Tucker notation,

Σ\Sigma1

For Σ\Sigma2 this reduces to the classical covariance Lyapunov equation; for Σ\Sigma3 one obtains the componentwise identity

Σ\Sigma4

(Recke et al., 17 Mar 2026).

The identifiability problem changes fundamentally when the noise is Gaussian. If Σ\Sigma5, then Σ\Sigma6 often cannot be recovered from Σ\Sigma7 alone in the unrestricted continuous model. Recke and Hansen therefore assume nonzero higher-order coordinate cumulants, as occurs for independent non-Gaussian jumps with diagonal Σ\Sigma8. Their main theorem states that if Σ\Sigma9 is a connected DAG with self-loops and MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.0, then outside a proper algebraic exception set the triple MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.1 is determined uniquely up to a common scalar factor by MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.2. Equivalently,

MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.3

has MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.4-dimensional generic fiber

MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.5

The residual ambiguity is the trivial time-rescaling MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.6, MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.7 (Recke et al., 17 Mar 2026).

The proof strategy vectorizes the second- and MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.8-th-order equations simultaneously, removes the rows and columns associated with diagonal cumulant parameters, and reduces the problem to a linear system in MΣ+ΣM⊤+C=0.M\Sigma+\Sigma M^\top + C = 0.9,

MM0

A special stable lower-triangular choice of MM1 together with generic diagonal cumulants yields rank MM2, and algebraic-geometric arguments then show that the one-dimensional kernel is generic (Recke et al., 17 Mar 2026).

A plausible implication is that acyclicity alone does not exhaust the identifiability structure of continuous Lyapunov models: once higher-order cumulants are observed, the same DAG restriction supports identification in broader non-Gaussian equilibrium models than the Gaussian covariance theory initially suggests.

5. Model equivalence and structural identifiability of DAGs

When the graph itself is unknown, a distinct question arises: which DAGs define the same acyclic GCLM? In the setting of stable drift matrices with uncorrelated noise

MM3

the acyclic GCLM of MM4 is

MM5

and two DAGs are model equivalent if MM6 (Améndola et al., 6 Oct 2025).

The first characterization is the MM7-node criterion: two DAGs MM8 on the same vertex set are model equivalent if and only if they have the same skeleton and, for every MM9-element subset M∈Rp×pM\in\mathbb R^{p\times p}00, the induced subgraphs M∈Rp×pM\in\mathbb R^{p\times p}01 and M∈Rp×pM\in\mathbb R^{p\times p}02 are themselves model equivalent (Améndola et al., 6 Oct 2025). This is the Lyapunov analogue of the Verma–Pearl characterization for Bayesian networks, but the local obstruction size increases from M∈Rp×pM\in\mathbb R^{p\times p}03 nodes to M∈Rp×pM\in\mathbb R^{p\times p}04.

The second characterization is transformational. An edge M∈Rp×pM\in\mathbb R^{p\times p}05 is super-covered if

  • M∈Rp×pM\in\mathbb R^{p\times p}06 and M∈Rp×pM\in\mathbb R^{p\times p}07,
  • every parent of M∈Rp×pM\in\mathbb R^{p\times p}08 is a parent of every child of M∈Rp×pM\in\mathbb R^{p\times p}09,
  • any third vertex M∈Rp×pM\in\mathbb R^{p\times p}10 is either a neighbor of both M∈Rp×pM\in\mathbb R^{p\times p}11 and M∈Rp×pM\in\mathbb R^{p\times p}12 or else has no trek connecting it to M∈Rp×pM\in\mathbb R^{p\times p}13 or to M∈Rp×pM\in\mathbb R^{p\times p}14.

Two simple DAGs satisfy M∈Rp×pM\in\mathbb R^{p\times p}15 if and only if one can transform M∈Rp×pM\in\mathbb R^{p\times p}16 into M∈Rp×pM\in\mathbb R^{p\times p}17 by a finite sequence of super-covered edge flips, and if they differ on M∈Rp×pM\in\mathbb R^{p\times p}18 edges then exactly M∈Rp×pM\in\mathbb R^{p\times p}19 flips suffice (Améndola et al., 6 Oct 2025).

These results have two important consequences. First, Lyapunov equivalence classes refine Bayesian-network equivalence classes: every pair of Lyapunov-equivalent DAGs is also Markov-equivalent as Bayesian networks, but not conversely. Second, both equivalence testing and structural identifiability admit polynomial-time algorithms. The equivalence test runs in M∈Rp×pM\in\mathbb R^{p\times p}20 by checking skeleton equality and all induced M∈Rp×pM\in\mathbb R^{p\times p}21-node subgraphs; structural identifiability can be tested in M∈Rp×pM\in\mathbb R^{p\times p}22 by searching for super-covered edges via the seven non-identifiable M∈Rp×pM\in\mathbb R^{p\times p}23-node patterns (Améndola et al., 6 Oct 2025). This corrects the common assumption that continuous-time DAG models inherit exactly the same equivalence theory as Gaussian Bayesian networks.

6. Estimation, model selection, and finite-sample behavior

Estimation methods for acyclic GCLMs split naturally into covariance-based Gaussian procedures and cumulant-based non-Gaussian procedures. In the non-Gaussian setting, Recke and Hansen propose a semiparametric moment-matching estimator. Given M∈Rp×pM\in\mathbb R^{p\times p}24 i.i.d. samples M∈Rp×pM\in\mathbb R^{p\times p}25, one computes empirical M∈Rp×pM\in\mathbb R^{p\times p}26 and M∈Rp×pM\in\mathbb R^{p\times p}27, forms the estimated coefficient matrix

M∈Rp×pM\in\mathbb R^{p\times p}28

and estimates M∈Rp×pM\in\mathbb R^{p\times p}29 by the right singular vector corresponding to the smallest singular value:

M∈Rp×pM\in\mathbb R^{p\times p}30

normalized by M∈Rp×pM\in\mathbb R^{p\times p}31 and the sign convention M∈Rp×pM\in\mathbb R^{p\times p}32. Under finite M∈Rp×pM\in\mathbb R^{p\times p}33-th moments,

M∈Rp×pM\in\mathbb R^{p\times p}34

so the estimator is M∈Rp×pM\in\mathbb R^{p\times p}35-consistent and asymptotically normal (Recke et al., 17 Mar 2026).

The same paper emphasizes the difficulty of the unconstrained estimation problem in finite samples. In simulations with M∈Rp×pM\in\mathbb R^{p\times p}36 up to approximately M∈Rp×pM\in\mathbb R^{p\times p}37 and compound-Poisson noise, the estimator showed noticeable bias unless M∈Rp×pM\in\mathbb R^{p\times p}38; the root-M∈Rp×pM\in\mathbb R^{p\times p}39-scaled bias decayed slowly when the noise correlation M∈Rp×pM\in\mathbb R^{p\times p}40 was large. To achieve M∈Rp×pM\in\mathbb R^{p\times p}41, sample sizes of order M∈Rp×pM\in\mathbb R^{p\times p}42–M∈Rp×pM\in\mathbb R^{p\times p}43 were often needed, with the requirement increasing in both dimension and correlation (Recke et al., 17 Mar 2026). Thus higher-order identifiability does not by itself imply easy estimation.

For Gaussian sparse learning, the Direct Lyapunov Lasso solves

M∈Rp×pM\in\mathbb R^{p\times p}44

or, after vectorization, a standard lasso with Hessian M∈Rp×pM\in\mathbb R^{p\times p}45. Under the irrepresentability condition

M∈Rp×pM\in\mathbb R^{p\times p}46

together with explicit sample-size and tuning-parameter inequalities, one obtains uniqueness, sup-norm error control, and exact support and sign recovery when the minimum signal exceeds the error bound (Dettling et al., 2022). The analysis also shows that this condition is unexpectedly hard to satisfy.

Acyclicity is central here as well. For DAG supports, if one starts at a diagonal drift M∈Rp×pM\in\mathbb R^{p\times p}47, then the irrepresentability condition holds uniformly in a neighborhood of M∈Rp×pM\in\mathbb R^{p\times p}48 precisely when

M∈Rp×pM\in\mathbb R^{p\times p}49

an ordering that is possible only for acyclic graphs (Dettling et al., 2022). In contrast, simulations on cyclic supports showed that the condition almost never holds. Nonetheless, numerical experiments found that the lasso recovered relevant structure robustly even under volatility misspecification, and on the Sachs et al. protein-signaling data a selected graph on dataset M∈Rp×pM\in\mathbb R^{p\times p}50 had M∈Rp×pM\in\mathbb R^{p\times p}51 directed edges, with roughly M∈Rp×pM\in\mathbb R^{p\times p}52–M∈Rp×pM\in\mathbb R^{p\times p}53 matching or reversing edges in the published network (Dettling et al., 2022).

The original structure-learning approach of Varando and Hansen instead minimizes a nonconvex penalized loss in M∈Rp×pM\in\mathbb R^{p\times p}54,

M∈Rp×pM\in\mathbb R^{p\times p}55

with either Gaussian negative log-likelihood or Frobenius loss, and uses a proximal-gradient algorithm in which each gradient step is computed through Lyapunov solves of cost M∈Rp×pM\in\mathbb R^{p\times p}56 (Varando et al., 2020). In historical terms, this optimization framework introduced sparse learning for GCLMs, while later work clarified when acyclicity yields exact covariance identifiability, when model equivalence remains nontrivial, and when higher-order cumulants can recover drift parameters beyond the Gaussian case.

Acyclic GCLMs therefore occupy a technically distinctive position among continuous-time graphical models: they are globally identifiable from covariance under fixed Gaussian noise, admit a refined equivalence theory that is stricter than Bayesian-network Markov equivalence, and remain identifiable in broader non-Gaussian settings through higher-order cumulants, while still presenting substantial finite-sample and model-selection challenges (Dettling et al., 2022, Améndola et al., 6 Oct 2025, Recke et al., 17 Mar 2026).

Topic to Video (Beta)

No one has generated a video about this topic yet.

Whiteboard

No one has generated a whiteboard explanation for this topic yet.

Follow Topic

Get notified by email when new papers are published related to Acyclic Graphical Continuous Lyapunov Models (GCLMs).