---
title: Noise-Induced Barren Plateaus in Quantum Circuits
url: https://www.emergentmind.com/topics/noise-induced-barren-plateaus-nibp
type: topic
---

# Noise-Induced Barren Plateaus in Quantum Circuits

Noise-Induced Barren Plateaus (NIBP) are a fundamental limitation to the scalability and trainability of variational quantum algorithms (VQAs) operating on noisy, pre-fault-tolerant hardware. NIBPs refer to the phenomenon where gradients of the cost function vanish exponentially due to noise accumulation, rendering classical parameter optimization intractable regardless of circuit initialization or ansatz structure. The interplay between noise model (especially unital vs non-unital channels), cost function locality, circuit connectivity, and algorithm design determines whether and when NIBPs manifest and whether they can be mitigated. Recent advances have revealed nuanced distinctions between noise types, effective circuit depth, and strategies for both diagnosing and overcoming NIBPs.

## 1. Theoretical Foundations: Definitions and Noise Models

Consider a parameterized quantum circuit with parameters $\theta=(\theta_1,\ldots,\theta_m)$ and cost function $C(\theta)=\operatorname{Tr}[H\,\Phi_\theta(\rho_0)]$, where $\Phi_\theta$ represents the full, possibly noisy evolution. A barren plateau occurs when the gradient variance, averaged over random parameter choices,
$$\mathbb{E}_\theta[\|\nabla C\|_2^2] = O(e^{-\Omega(n)})$$
is exponentially suppressed in the number of qubits $n$ (or, equivalently, $\,\mathrm{Var}_\theta[\partial_i C]=E_\theta[(\partial_i C)^2]$ is exponentially small for each parameter $i$).

**Noise-Induced Barren Plateaus (NIBP)** specifically arise when the decay is induced by noise channels rather than ansatz expressivity or cost function structure. For general noise, including local Pauli, depolarizing, or more general completely positive trace-preserving (CPTP) maps, the circuit is modeled as layers of local gates, each followed by a noise channel $N$, described in a "normal form" as
$$
N[(I + w\cdot\sigma)/2]= I/2 + \frac12 (t + D w)\cdot\sigma,
$$
with $t$ (affine shift) and $D$ (contraction) determining unital ($t=0$, e.g., depolarizing) and non-unital ($t\neq 0$, e.g., amplitude damping) regimes [2403.13927].

Unital noise channels (e.g., depolarizing, dephasing) preserve the maximally mixed state and exponentially drive all cost function gradients—and expectation values for traceless observables—toward zero as depth increases [2402.08721, 2007.14384, 2405.00781]. Non-unital channels (e.g., amplitude damping, reset) have a nontrivial affine term ($t\neq 0$), which can preserve non-zero gradients in certain regimes.

## 2. NIBP Mechanisms and Gradient-Variance Suppression

Under realistic noise, the gradient variance associated with any trainable parameter decays as
$$
\mathrm{Var}_\theta[\partial_\mu C] \leq O\left(c^{|P| + L - k - 1}\right)
$$
for a Pauli observable $P$ of weight $|P|$ in an $L$-layer circuit, with $k$ the layer index and $c<1$ a contraction factor determined by the noise properties [2403.13927]. When $N$ is unital, $c$ is strictly less than 1, and exponential suppression dominates for any global cost function and at all layers:
$$
\mathrm{Var}_\theta[\partial_i C] = O(\alpha^L/2^n)
$$
with $\alpha < 1$ and $L$ large, implying an NIBP for any circuit whose depth grows with $n$ [2405.00781, 2007.14384, 2310.08405].

For local cost functions and non-unital noise, this gradient variance is exponentially suppressed except in the last $O(\log n)$ layers, where it can remain polynomially large:
$$
\mathrm{Var}_\theta[\partial_\mu C] \geq \exp(-O(|P| (L - k)))
$$
if $N$ is non-unital and $H_\mu$ acts within the light cone of $P$ [2403.13927]. Thus, for local observables, only the last few layers remain trainable; the rest of the circuit is effectively "frozen" and does not affect the cost landscape (“effective shallowness”).

For global cost observables (e.g., full state infidelity), all gradients vanish exponentially for both unital and non-unital noise, guaranteeing a NIBP irrespective of noise type or ansatz [2403.13927, 2402.08721].

## 3. Effective Circuit Shallowness and Classical Simulability

Noise does not merely induce gradient suppression; it also "truncates" the effective quantum circuit depth. More precisely, the expectation of a local Pauli $P$ after a depth-$L$ circuit under noise is, up to an exponentially small error, determined solely by the final $m=O(\log(1/\epsilon))$ layers:
$$
|\operatorname{Tr}(P\,\Phi(\rho)) - \operatorname{Tr}(P\,\Phi_{[L-m,L]}(\sigma))| \leq \|P\|_\infty c^{m+|P|-1} \leq \epsilon
$$
[2403.13927]. The circuit becomes "effectively shallow" for all observable estimation tasks: for local cost functions and unital or non-unital noise, only the last $O(\log n)$ layers affect measurable outcomes or gradients.

This effective shallowness has a direct impact on classical simulation complexity. By propagating observables backward through only the last $O(\log n)$ layers, classical algorithms can estimate expectation values with runtime polynomial in $n$ for 1D circuits and quasi-polynomial for higher-dimensional connectivity, irrespective of total physical depth [2403.13927]. This principle underpins efficient classical simulation and rules out quantum advantage for expectation value estimation under generic noise without error correction.

## 4. Experimental Observations and Absence of NIBP under Non-Unital Noise

Experimental studies using large-scale superconducting hardware (IBM Falcon and Heron processors, up to 102 qubits) have tested NIBP predictions in regimes dominated by amplitude damping (non-unital, $T_1$ relaxation) noise versus depolarizing (unital) noise [2602.22851]. Measurement of the average gradient norm versus circuit runtime using Information Content Landscape Analysis (ICLA) reveals:

- Under depolarizing noise, gradient norms decay exponentially to zero within characteristic circuit times, consistent with NIBP theory.
- Under amplitude damping, gradient norms saturate to a finite plateau beyond a hardware- and noise-dependent "flattening time" $t_\text{flat}$. There is no exponential vanishing, and finite gradient magnitudes persist at long circuit depths, even up to $N=102$ qubits.
- The quantum hardware's effective coherence time $T_1^\mathrm{eff}\approx 0.72\,t_\text{flat}$ is determined by the worst-performing qubits, not by the average $T_1$ calibration value. Device benchmarking based solely on average metrics may underestimate the onset and severity of trainability bottlenecks [2602.22851].

Classical simulations corroborate experimental findings: NIBPs are observed under depolarizing noise but are absent under amplitude damping, due to the fixed-point structure and residual parameter dependence in the steady state for non-unital channels. Local cost function optimization remains feasible in the presence of non-unital noise.

## 5. Mitigation and Avoidance of NIBP: Dissipative Algorithms and Engineered Noise

Several strategies have emerged for mitigating or circumventing NIBP:

**1. Engineered Dissipation and Non-Unital Circuits:**  
Embedding dissipative steps or periodic resets within the circuit introduces non-unital channels that can "extract entropy," maintain nonzero gradients, and guarantee trainability even at large depths [2310.15037, 2507.02043]. For example, resetting a fraction $n_r/n$ of ancillary qubits via amplitude damping after every $L=O(\log n)$ layers delivers a lower bound
$$
\mathrm{Var}[\partial_\mu C] \geq \Omega(1/\mathrm{poly}(n))
$$
for parameters in the light cone of the observable, eliminating NIBP for local observables. Analytic conditions and numerical evidence confirm scalable trainability with this approach, even as the circuit depth increases [2507.02043].

**2. Non-Unitary Variational Ansätze:**  
Introduced in both mean-field models and realistic quantum chemistry simulations, incorporating jump operators and non-unitary dynamics directly into the variational layers (i.e., variational Lindblad channels) enables preparation of open-system steady states and preservation of trainability under realistic noise [2605.30572]. Analysis of circuit fixed points reveals that while purely unitary, unital circuits converge to the maximally mixed state (and NIBP), non-unitary channels with multiple steady states result in optimization landscapes retaining nontrivial, parameter-dependent structure.

**3. Local Pre-Training Strategies:**  
For applications such as geometric entanglement measurement, sequentially optimizing a series of commuting local cost functions, followed by global cost refinement, can enable escape from noise-induced barren plateaus without needing non-unitary channels [2304.13388].

## 6. Practical Implications and Remaining Challenges

The discovery and analysis of NIBPs have major implications:

- **Trainability bounds:** In the absence of error correction or engineered dissipation, any variational quantum circuit suffering unital noise will manifest an NIBP when the depth scales with system size. Only circuits with constant/logarithmic depth, or circuits executing with non-unital noise channels, remain trainable for local observables [2403.13927, 2402.08721].
- **Limitations for quantum advantage:** For algorithms whose outputs are local observable estimations (including many quantum machine learning and quantum chemistry protocols), noise-induced effective shallowness and NIBP preclude scaling advantages without fault tolerance or dissipative engineering. Only shallow circuits or hybrid quantum-classical strategies may remain viable.
- **Device benchmarking:** Reliable prediction of NIBP onset requires statistical analysis of device noise distributions, not mere averages; performance is dominated by the worst-case qubits and coherence times [2602.22851].
- **Generalization beyond unital noise:** Rigorous results confirm that NIBP only occurs generically in circuits with unital noise. Non-unital (Hilbert-Schmidt contractive) noise yields noise-induced limit sets (NILS), in which cost values concentrate within a parameter-dependent interval but do not uniformly collapse gradients to zero [2402.08721].
- **Open questions:** Effective-depth bounds for worst-case circuits, architectural evasion mechanisms, complexity of sampling from noisy circuits, and connection to measurement-induced phase transitions remain open research areas [2403.13927].

## 7. Summary Table: NIBP Behavior by Noise Model

| Noise Model         | Gradient Variance Decay         | Trainability for Local Cost | Global Cost/Observable |
|---------------------|---------------------------------|----------------------------|-----------------------|
| Unital (e.g., depol, dephasing)   | Exponential in depth/system size | No (\textit{NIBP})             | Exponentially flat    |
| Non-unital (e.g., amplitude damping, engineered reset) | No exponential decay in last $O(\log n)$ layers | Yes (locally, only last layers trainable) | Still untrainable    |

*Table summarizes critical findings from [2403.13927, 2007.14384, 2402.08721, 2602.22851, 2507.02043].*

## References

- "Noise-induced shallow circuits and absence of barren plateaus" [2403.13927]
- "Experimental demonstration of the absence of noise-induced barren plateaus using information content landscape analysis" [2602.22851]
- "Barren Plateaus in Variational Quantum Computing" [2405.00781]
- "Mitigating Noise-Induced Barren Plateaus Using a Non-Unitary Ansatz" [2605.30572]
- "Engineered dissipation to mitigate barren plateaus" [2310.15037]
- "Scaling Quantum Algorithms via Dissipation: Avoiding Barren Plateaus" [2507.02043]
- "Emergence of noise-induced barren plateaus in arbitrary layered noise models" [2310.08405]
- "Avoiding barren plateaus in the variational determination of geometric entanglement" [2304.13388]
- "Noise-Induced Barren Plateaus in Variational Quantum Algorithms" [2007.14384]
- "Beyond unital noise in variational quantum algorithms: noise-induced barren plateaus and limit sets" [2402.08721]

Source: https://www.emergentmind.com/topics/noise-induced-barren-plateaus-nibp