Papers
Topics
Authors
Recent
Search
2000 character limit reached

Small-Covariance Noise-to-State Stability (scNSS)

Updated 14 July 2026
  • scNSS is a stochastic stability concept that ensures a system’s state remains within a covariance-dependent neighborhood of equilibrium when noise is sufficiently small.
  • It utilizes Lyapunov-based inequalities with bounded covariance gains to distinguish between global NSS, local scNSS, and integral NSS frameworks.
  • Key applications include stochastic gradient dynamics, LQR, and logistic regression, where controlling noise covariance is critical for desired stability.

Small-covariance noise-to-state stability (scNSS) is a stochastic stability notion for systems whose state remains, with high probability, in a covariance-dependent neighborhood of an equilibrium, but only when the driving noise covariance is sufficiently small. In the formulation introduced for stochastic gradient dynamics, scNSS preserves the standard NSS template

P ⁣{V(χ(t))β(V(χ(0)),t)+γ(ΣΣ)}1ϵ,\mathbb P\!\left\{ \mathcal V(\chi(t))\le \beta(\mathcal V(\chi(0)),t)+\gamma(\Sigma\Sigma^\top) \right\}\ge 1-\epsilon,

while restricting the admissible covariance magnitude to

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,

with γK[0,d)\gamma\in\mathcal K_{[0,d)} rather than γK\gamma\in\mathcal K. The concept was introduced to capture robustness regimes in which stability is available only for sufficiently small covariance, especially in stochastic optimization and Langevin-type dynamics (Cui et al., 29 Sep 2025).

1. Formal definition and state-space setting

The scNSS framework is posed for stochastic systems of the form

dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),

where χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n, S\mathcal S is open and diffeomorphic to Rn\mathbb R^n, B(t)B(t) is an mm-dimensional standard Brownian motion, ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,0 and ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,1 are locally bounded and locally Lipschitz, ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,2 at an equilibrium ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,3, and ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,4 is Borel measurable and locally essentially bounded. The instantaneous noise covariance is ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,5 (Cui et al., 29 Sep 2025).

A size function ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,6 is twice continuously differentiable, positive definite with respect to ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,7, and coercive in the sense that

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,8

This function supplies the state-energy variable in the stability estimate.

The distinction between NSS and scNSS is entirely in the disturbance channel. Standard NSS in probability requires the bound

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,9

for some γK[0,d)\gamma\in\mathcal K_{[0,d)}0, γK[0,d)\gamma\in\mathcal K_{[0,d)}1, all γK[0,d)\gamma\in\mathcal K_{[0,d)}2, all initial conditions, and all γK[0,d)\gamma\in\mathcal K_{[0,d)}3. By contrast, scNSS in probability uses the same probabilistic form but only for sufficiently small covariance, with γK[0,d)\gamma\in\mathcal K_{[0,d)}4 and the side condition γK[0,d)\gamma\in\mathcal K_{[0,d)}5. The definition therefore encodes a local-in-covariance robustness property rather than a global-in-covariance one. The same source also recalls integral NSS (iNSS),

γK[0,d)\gamma\in\mathcal K_{[0,d)}6

which weakens the disturbance channel further by allowing an accumulated energy term instead of an instantaneous covariance gain.

2. Lyapunov characterization and hierarchy of stochastic robustness

The central Lyapunov inequalities separate NSS, scNSS, and iNSS by the admissible comparison function class in the dissipation term. An NSS-Lyapunov function satisfies

γK[0,d)\gamma\in\mathcal K_{[0,d)}7

for γK[0,d)\gamma\in\mathcal K_{[0,d)}8, γK[0,d)\gamma\in\mathcal K_{[0,d)}9, all γK\gamma\in\mathcal K0, and all γK\gamma\in\mathcal K1. An scNSS-Lyapunov function satisfies the same inequality with γK\gamma\in\mathcal K2, γK\gamma\in\mathcal K3, and only for γK\gamma\in\mathcal K4. An iNSS-Lyapunov function weakens dissipation further to γK\gamma\in\mathcal K5, γK\gamma\in\mathcal K6 (Cui et al., 29 Sep 2025).

The corresponding implication results are direct: existence of an scNSS-Lyapunov function implies scNSS, existence of an NSS-Lyapunov function implies NSS, and existence of an iNSS-Lyapunov function implies iNSS. In the scNSS proof, stopping times and supermartingale arguments are used to show that trajectories enter and remain in a neighborhood

γK\gamma\in\mathcal K7

whose size is controlled by the covariance. A representative estimate is

γK\gamma\in\mathcal K8

valid when γK\gamma\in\mathcal K9 is below a small-covariance threshold dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),0.

Within this framework, NSS is the strongest notion, scNSS is weaker because it only covers dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),1, and iNSS is weaker still because it only requires positive-definite dissipation and admits an integral disturbance term. The same hierarchy appears at the level of optimization geometry: dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),2 A common misconception is to treat scNSS as merely NSS with smaller constants. The formal difference is sharper: the gain itself is only defined on a bounded covariance interval, so the guarantee is structurally local in covariance magnitude.

3. Gradient dynamics and generalized Polyak–Łojasiewicz structure

The principal application domain for scNSS is stochastic gradient dynamics of the overdamped Langevin type,

dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),3

The standing assumptions are that dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),4, dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),5 is bounded below and coercive, and dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),6 is locally Lipschitz and locally bounded. When dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),7 is globally dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),8-Lipschitz and dχ(t)=f(χ(t))dt+g(χ(t))Σ(t)dB(t),d\chi(t)=f(\chi(t))\,dt+g(\chi(t))\,\Sigma(t)\,dB(t),9 is globally bounded by χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n0, the Lyapunov candidate

χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n1

satisfies

χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n2

This makes the generalized Polyak–Łojasiewicz condition the decisive geometric hypothesis (Cui et al., 29 Sep 2025).

The paper distinguishes three variants: χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n3 with χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n4 for χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n5-PL, χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n6 for χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n7-PL, and χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n8 for positive-definite-PL. Under global Lipschitzness and bounded noise gain, the theorem for overdamped diffusion states that χ(t)SRn\chi(t)\in\mathcal S\subset\mathbb R^n9-PL implies NSS, S\mathcal S0-PL implies scNSS and iNSS, and S\mathcal S1-PL implies iNSS. The neighborhood size around the optimum is determined by the covariance term in the Lyapunov inequality; smaller covariance produces a smaller ultimate region.

The same trichotomy extends to underdamped Langevin diffusion, or stochastic heavy-ball dynamics,

S\mathcal S2

Here the analysis uses the mixed Lyapunov function

S\mathcal S3

and obtains an estimate of the form

S\mathcal S4

Under global Lipschitz gradient and bounded S\mathcal S5, the same three-way conclusion holds: S\mathcal S6-PL yields NSS, S\mathcal S7-PL yields scNSS and iNSS, and S\mathcal S8-PL yields iNSS.

4. Non-globally-Lipschitz objectives and step-size tuning

When S\mathcal S9 is not globally Lipschitz, the fixed-drift overdamped model is replaced by a state-dependent step-size dynamics,

Rn\mathbb R^n0

The analysis is organized through sublevel sets

Rn\mathbb R^n1

and curvature envelopes

Rn\mathbb R^n2

Under the Rn\mathbb R^n3-PL condition, the ratio

Rn\mathbb R^n4

determines whether the system is merely small-covariance stable or globally NSS (Cui et al., 29 Sep 2025).

The resulting dichotomy is explicit. If

Rn\mathbb R^n5

then the dynamics are scNSS. If instead

Rn\mathbb R^n6

the dynamics are NSS. The interpretation given is that the learning rate Rn\mathbb R^n7 must dominate the local curvature growth Rn\mathbb R^n8 relative to the PL gain. Two concrete examples are provided: Rn\mathbb R^n9 gives scNSS, whereas B(t)B(t)0 gives NSS.

For the underdamped non-globally-Lipschitz case, a different Lyapunov construction is used,

B(t)B(t)1

where B(t)B(t)2 is obtained by smoothing a B(t)B(t)3 bound on the gradient via local curvature. Under suitable tuning of B(t)B(t)4 and B(t)B(t)5, the same B(t)B(t)6 trichotomy again yields NSS, scNSS, and iNSS. This suggests that scNSS is not tied to global smoothness, but rather to the balance among curvature growth, dissipation, and covariance magnitude.

5. Canonical applications: LQR and logistic regression

For continuous-time linear quadratic regulator policy optimization, the objective is

B(t)B(t)7

where B(t)B(t)8 solves

B(t)B(t)9

and the gradient is

mm0

with mm1 solving

mm2

The objective mm3 is coercive and analytic, mm4 is only locally Lipschitz, and mm5 satisfies the mm6-PL condition

mm7

For the overdamped stochastic policy dynamics,

mm8

the criterion is

mm9

while

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,00

For underdamped or heavy-ball LQR policy dynamics, the same general underdamped theorem and the ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,01-PL property yield scNSS.

For logistic regression, the loss is

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,02

Under the assumptions that the data matrix ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,03 has full row rank and the data are nonseparable, ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,04 is coercive, strictly convex, and has globally Lipschitz gradient with constant

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,05

It also satisfies a ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,06-PL condition

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,07

for some ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,08-function ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,09, but it does not satisfy the ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,10-PL condition because ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,11 is uniformly bounded while ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,12 as ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,13. Consequently, both the overdamped diffusion

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,14

and the underdamped dynamics

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,15

are concluded to be scNSS rather than NSS. The operational interpretation is that the stochastic training trajectory remains, with high probability, in a covariance-dependent neighborhood of the logistic optimum, and that neighborhood shrinks as the covariance decreases (Cui et al., 29 Sep 2025).

6. Relation to earlier NSS theory and adjacent extensions

scNSS sits within a broader NSS lineage for stochastic systems with persistent or state-dependent noise. Earlier NSS theory for SDEs with persistent noise already expressed state bounds in terms of the maximum covariance magnitude through ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,16, using NSS in probability and ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,17th-moment NSS with Lyapunov conditions of the form

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,18

That framework did not isolate a separate small-covariance notion, but it already encoded the principle that smaller covariance implies a smaller disturbance-dependent state bound (Mateos-Núñez et al., 2015).

A different precursor is the time-domain analysis of noise-to-state exponentially stable systems. For SDEs

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,19

the resident-time ratio

ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,20

was introduced to quantify the fraction of time a trajectory spends in a bounded region. The resulting almost-sure lower bound improves as the noise intensity decreases and tends to ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,21 as the noise vanishes. That work does not study a small-covariance regime in the perturbative sense, but it provides a time-average small-noise enhancement of NSS-type behavior and complements fixed-time probabilistic boundedness by a long-run residence guarantee (Fang et al., 2016).

More recent neighboring developments broaden the same theme rather than replacing it. Stochastic contraction results give incremental noise- and input-to-state stability in weighted ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,22-norms, with asymptotic mean-square errors scaling linearly with diffusion magnitude, such as terms of order ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,23 and ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,24; this is closely aligned with the intuition behind small-covariance robustness (Kawano et al., 20 Feb 2026). In Wasserstein space, distributional ISS has been proposed for measure-valued dynamics, with the claim that on compact domains it recovers particle-level ISS and NSS, including stochastic particle systems whose disturbance magnitude is ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,25 (Pascual et al., 30 Mar 2026). For discrete-time linear systems, risk-aware stability introduces NSS-type offsets measured through coherent risk functionals or a mean-conditional-variance functional, so that covariance and higher-order centered noise statistics affect the residual state-risk bound (Chapman et al., 2022). In randomly switching linear systems with additive Gaussian noise, lifted covariance dynamics and invariant ellipsoids yield computable bounds on state covariance trajectories; this has been interpreted as a small-covariance NSS-type property because smaller mode-dependent covariances ΣΣ<d,\|\Sigma\Sigma^\top\|_\infty<d,26 lead to tighter covariance confinement (Yoon et al., 2019).

Taken together, these results delimit the scope of scNSS. It is neither an exact stationary-distribution theory nor a generic ergodic theorem, and it is distinct from fixed-time NSS, time-average residence bounds, distributional Wasserstein robustness, and risk-aware formulations. Its specific role is to formalize the regime in which covariance-dependent ultimate boundedness is valid only below a noise threshold, a setting that arises naturally in stochastic gradient systems and related control and optimization dynamics.

Topic to Video (Beta)

No one has generated a video about this topic yet.

Whiteboard

No one has generated a whiteboard explanation for this topic yet.

Follow Topic

Get notified by email when new papers are published related to Small-Covariance Noise-to-State Stability (NSS).