---
title: Local Asymptotic Normality (LAN)
url: https://www.emergentmind.com/topics/local-asymptotic-normality-lan
type: topic
---

# Local Asymptotic Normality (LAN)

Local Asymptotic Normality (LAN) is a foundational property in statistical inference, characterizing the asymptotic behavior of likelihood ratios in families of stochastic processes or statistical experiments. Informally, a sequence of statistical models exhibits LAN if, under suitable local reparameterization near a "true" parameter value, the (log-)likelihood ratio between the perturbed and true parameters admits a universal Gaussian quadratic approximation. This property is pivotal in asymptotic decision theory, underlies the construction of efficient estimators, and determines optimal testing strategies. The LAN property has been established—often with considerable technical sophistication—for a broad class of stochastic models, including but not limited to interacting diffusions, jump processes, fractional and mixed Brownian models, nonparametric drift estimation, multi-armed bandits, and quantum statistical models.

## 1. Definition and General Formalism

Let \((P_\vartheta^n)\) denote a sequence of statistical models, parametrized by \(\vartheta\) in an open subset \(\Theta\) of \(\mathbb{R}^p\) or more generally in a separable Hilbert space. The sequence is said to have the **Local Asymptotic Normality (LAN)** property at \(\vartheta_0\) if there exist:

- A scaling or normalization matrix \(\varphi_n(\vartheta_0)\), typically \(\varphi_n = n^{-1/2} I_p\) in regular parametric models;
- A random **central sequence** (or local score) \(\Delta_n(\vartheta_0)\);
- A symmetric positive-definite **Fisher information matrix** \(I(\vartheta_0)\);

such that for every fixed \(u \in \mathbb{R}^p\),
\[
\log \frac{dP^n_{\vartheta_0 + \varphi_n(\vartheta_0)u}}{dP^n_{\vartheta_0}} = u^\top \Delta_n(\vartheta_0) - \frac{1}{2} u^\top I(\vartheta_0) u + o_{P^n_{\vartheta_0}}(1),
\]
and under \(P^n_{\vartheta_0}\), \(\Delta_n(\vartheta_0) \xrightarrow{d} N(0,\, I(\vartheta_0))\). All classical asymptotic theory for efficient estimation (e.g., Cramér–Rao lower bounds, construction of locally most powerful tests, LAN-based minimax theorems) follows from this structure [2205.05932, 2512.12192, 2007.00723].

In nonstandard or high-frequency settings the normalization \(\varphi_n\) may be non-diagonal or process-dependent, reflecting different estimation rates for different parameters [1610.03694, 2512.24042, 2601.02622].

## 2. Methodologies for Establishing LAN

The proof of the LAN property is model-dependent but always relies on a combination of the following elements:

- **Taylor Expansion of Log-Likelihood:** Expansion in local parameter neighborhoods, isolating linear (score) and quadratic (Fisher information) terms.
- **Martingale and Central Limit Theorems:** Martingale CLT for the central sequence or its projections is routinely used in models involving stochastic differential equations or Markov processes [2511.13366, 1509.00003, 2205.05932].
- **Girsanov/Change-of-Measure Formulas:** Girsanov’s theorem for continuous and jump diffusions supplies explicit formulas for likelihood ratios [2205.05932, 1509.00003, 1402.4956].
- **Spectral and Semigroup Methods:** In nonparametric and time-homogeneous settings (e.g., for reflected diffusions), analytic/spectral PDE tools control the statistical random terms [1802.02009].
- **Malliavin Calculus & Integration by Parts:** In models with jumps or implicitly defined transition densities (e.g., McKean–Vlasov diffusions), Malliavin calculus enables explicit score representations [2511.13366, 1903.00358].
- **Operator and Non-diagonal Normalizations:** In time-series with strong dependency, e.g., fractional Gaussian noise or mixed fractional Brownian motion, orthogonalizations and rate matrices aligned to the degeneracy/singularity structure are necessary [2512.24042, 2601.02622, 1610.03694].
- **Quantum Statistical Models:** LAN extends to finite-dimensional quantum models, using noncommutative Lebesgue decompositions and quantum central limit theory [1703.07535, 1210.3749].

Special care is needed in non-ergodic, singular, or ill-posed statistical models, where the LAN expansion may fail, or the central sequence and Fisher information may be degenerate or require careful re-normalization.

## 3. LAN in Interacting Particle Systems and McKean-Vlasov SDEs

For mean-field models (McKean–Vlasov SDEs), consider N exchangeable particles:
\[
dX^i_t = b(\vartheta; t, X^i_t, \mu^N_t)dt + \sigma(t, X^i_t)dB^i_t, \quad \mu^N_t = N^{-1} \sum_{j=1}^N \delta_{X^j_t}
\]
The LAN property for fixed time horizon \(T\) and high-dimensional \(X^{(N)}\) is established via a continuous-time Girsanov formula, expanding the log-likelihood for a parameter shift \(\vartheta = \vartheta_0 + N^{-1/2}u\):
\[
\ell^N(\vartheta) - \ell^N(\vartheta_0) = u^\top \Delta_N(\vartheta_0) - \tfrac{1}{2}u^\top I(\vartheta_0)u + o_{P^N_{\vartheta_0}}(1),
\]
where
\[
\Delta_N(\vartheta_0) = \frac{1}{\sqrt N}\sum_{i=1}^N \int_0^T \nabla_{\vartheta}(c^{-1/2}b)(\vartheta_0; t, X^i_t, \mu^N_t)^\top dB^{i,N}_{t}
\]
(the empirical score) and
\[
I(\vartheta_0) = \lim_{N \to \infty} I_N(\vartheta_0) = \text{asymptotic Fisher information matrix}.
\]
Key technical requirements include Lipschitz continuity and regularity in \(b\) and \(\sigma\), uniform ellipticity, and injectivity/nondegeneracy of the Fisher information [2205.05932]. In discrete and high-frequency settings, similar LAN expansions hold but with possibly distinct scaling rates for drift and diffusion [2511.13366].

## 4. LAN in Models with Nonstandard Rates and Degeneracies

### Fractional Gaussian Models

In models involving fractional Brownian motion or its increments (fGn), standard diagonal normalization fails due to dependency-induced co-linearity in score vectors. Non-diagonal rate matrices must be constructed:
- For fGn with unknown Hurst parameter \(H\) and scale \(\sigma\):
  \[
  \varphi_n = \frac{1}{\sqrt n}\begin{pmatrix} 1 & 0 \\ -\sigma \log \Delta_n & \sigma \end{pmatrix}
  \]
  This "rotates" the raw score, yielding asymptotically independent components and non-singular Fisher information [1610.03694].
- For mixed fBm models \(Y_t = \sigma B^H_t + W_t\), further triangular or orthogonalizing transformations are necessary to address degeneracy, especially when \(H > 3/4\), ensuring the LAN expansion holds with a full-rank (possibly diagonalized) information matrix [2512.24042, 2601.02622].

### Multi-Component Parameters and Partial Information

In bandit models and adaptive allocation designs, LAN rates may be componentwise: parameters affecting optimal arms are \(\sqrt{T}\)-estimable, while those associated solely with suboptimal arms are only estimable at a \((\log T)\)-rate, and a block-diagonal information structure arises [2512.12192].

## 5. Infinite-Dimensional and Nonparametric LAN

LAN theory generalizes to infinite-dimensional parameter spaces, including nonparametric drift estimation in diffusions and convex M-estimation in Hilbert spaces:
- **Nonparametric Drift [1802.02009]:** The score is a directional derivative, and the Fisher information is an operator norm. PDE and spectral methods establish smoothness and rates.
- **Convex M-Estimation [1704.02840]:** With Mosco-convergence of objective functions (weaker than uniform convergence), LAN holds in the Hilbert space: for \(t \in \mathcal{H}\),
  \[
  n[F_n(\theta_0 + t/\sqrt n) - F_n(\theta_0)] \to \langle t, W \rangle + \tfrac{1}{2} \langle V t, t \rangle
  \]
  where \(W \sim N(0,A)\) and \(V\) is a generalized Hessian.

## 6. Minimax Theory, Efficiency, and Statistical Consequences

The principal consequence of LAN is that the maximum likelihood estimator (MLE), or any regular estimator whose asymptotic behavior matches the LAN quadratic expansion, achieves the minimax lower bound—the so-called convolution bound of Hájek and Le Cam—for local (contiguous) parameter neighborhoods:
\[
\liminf_{n\to\infty} \sup_{\|\vartheta' - \vartheta\| \leq \delta / \sqrt n} \mathbb{E}_{\vartheta'}\left[ w\left(\sqrt{n}(I(\vartheta)^{1/2}(\hat{\vartheta}_n - \vartheta'))\right) \right] \geq \int_{\mathbb{R}^p} w(z) \frac{1}{(2\pi)^{p/2}}e^{-|z|^2/2} dz,
\]
with equality for the asymptotically efficient estimator [2205.05932, 2512.24042, 2605.03698]. This encompasses the construction of efficient confidence regions, tests, and reduces local inference to the canonical Gaussian shift experiment. Analogous results hold in quantum statistical models, with the optimal estimation rates characterized via the quantum Fisher information and the Holevo bound [1703.07535, 1210.3749].

## 7. Extensions, Variations, and Related Notions

Several notable generalizations and variants are directly linked to the LAN paradigm:

- **Local Asymptotic Quadraticity (LAQ) and Mixed Normality (LAMN):** When the Gaussian approximation holds only in a weaker or random-coefficient sense (e.g., supercritical jump CIR, critical bandits), the LAN property is replaced by LAQ or LAMN, and limit likelihood ratios may involve more complex, possibly process-dependent, limit experiments [1903.00358].
- **Non-linear and Mean-Field Regimes:** In interactive particle systems or non-linear McKean–Vlasov models, identifiability and nondegeneracy of the Fisher information may require explicit analytic criteria, often computable via the moments of the limiting empirical measure or propagation-of-chaos principles [2205.05932].
- **Rescaled LAN (RLAN):** For models with growing Monte Carlo error or in large neighborhoods, cubic approximations extend the LAN notion (RLAN), and modified estimators (e.g., maximum cubic likelihood estimator) retain statistical efficiency even as traditional LAN error rates become negligible [2007.00723].

## Table: Representative Models and Features of LAN

| Model/Class                         | Normalization/Rate              | Key Techniques         |
|-------------------------------------|---------------------------------|-----------------------|
| Classical parametric (iid)          | \(\varphi_n = n^{-1/2} I_p\)   | Score / Fisher info   |
| McKean–Vlasov SDEs (continuous)     | \(N^{-1/2}\)                    | Girsanov, CLT         |
| High-frequency fGn / mixed fBm      | Non-diagonal, componentwise     | Orthogonalization     |
| Multi-armed bandits                 | Block-diagonal, componentwise   | Martingale CLT        |
| Nonparametric (diffusion drift)     | Local path in Hilbert space     | Spectral/PDE methods  |
| Quantum parametric (finite-dim)     | \(n^{-1/2}\), operator log-like | CCR, SLD, QCLT        |

LAN is thus the structural backbone of high-dimensional, high-frequency, and non-standard inference for both classical and quantum models, enabling sharp efficiency, testing, and information-theoretic bounds across a wide scope of modern statistical theory [2205.05932, 2512.24042, 1610.03694, 2007.00723, 2511.13366, 2512.12192, 1703.07535, 1210.3749].

Source: https://www.emergentmind.com/topics/local-asymptotic-normality-lan