---
title: Belavkin Equations in Quantum Filtering
url: https://www.emergentmind.com/topics/belavkin-equations
type: topic
---

# Belavkin Equations in Quantum Filtering

Searching arXiv for the cited Belavkin-equation papers to ground the article in current arXiv records.
arxiv_search query: "Belavkin equation quantum filtering coherent states squeezed light mean-field repeated measurements"
Belavkin equations are stochastic evolution equations for conditioned quantum states under continuous observation. In the standard non-demolition setting, one observes a commutative output process generated by an open quantum system coupled to bosonic fields, and the aim is to compute the conditional expectation of a Heisenberg observable \(X\) given the observation history up to time \(t\). In this sense, the Belavkin–Kushner–Stratonovich equation is the quantum analogue of nonlinear filtering, while its linear unnormalized counterpart is the quantum analogue of a Zakai equation [1303.0578]. A complementary discrete viewpoint derives the same equations as scaling limits of repeated indirect measurements and iterated Bayesian updates, in either a Brownian diffusive limit or a Poissonian jumpy limit [1210.0425].

## 1. Conditional-state dynamics and the non-demolition setting

The basic object is the conditional estimate
\[
\pi_t(X)=\mathbb{E}\big[j_t(X)\mid \mathfrak{Y}^{\mathrm{out}}_t\big],
\]
where \(j_t(X)=V(t)^*[X\otimes I]V(t)\) is the Heisenberg-evolved observable and \(\mathfrak{Y}^{\mathrm{out}}_t\) is the commutative von Neumann algebra generated by the observation record up to time \(t\) [1303.0578]. The non-demolition property,
\[
j_t(X)\in (\mathfrak{Y}^{\mathrm{out}}_t)',
\]
is the structural condition that makes conditioning meaningful: evolved system observables commute with all past observations [1303.0578].

In the Hudson–Parthasarathy framework, the system is coupled to bosonic input fields, and the measurement is performed on an output process obtained after interaction. Measured outputs may be quadratures, photon counts, or more general commuting linear combinations of output fields. In the phase-space treatment of multichannel diffusive measurements, the measured process is written as \(Z=FY\), with
\[
FF^T\succ 0,\qquad FJF^T=0,
\]
so that the observation channels are self-commuting and therefore classical [1602.07911].

The same filtering structure emerges from repeated indirect measurements. There, one starts from a discrete Bayesian recursion
\[
Q_n(\alpha| i_1,\cdots,i_n)= \frac{p(i_n|\alpha)\, Q_{n-1}(\alpha| i_1,\cdots,i_{n-1})} {\pi_{n-1}(i_n| i_1,\cdots,i_{n-1})},
\]
for pointer-state probabilities, and then passes to a continuous-time limit. In the quantum mechanical framework, this continuous time limit leads to Belavkin equations [1210.0425]. This suggests that the stochastic terms in Belavkin equations are best viewed as continuum innovations extracted from measurement records.

## 2. Hudson–Parthasarathy structure and derivational methods

For an open Markov quantum system with system Hilbert space \(\mathfrak h\), the unitary cocycle \(V(t)\) is governed by the Hudson–Parthasarathy QSDE
\[
dV(t)=\Big\{(S-I)\otimes d\Lambda(t) +L\otimes dB^*(t) -L^*S\otimes dB(t) -\Big(\frac12L^*L+iH\Big)\otimes dt\Big\}V(t),
\]
with unitary scattering \(S\), coupling operator \(L\), and self-adjoint Hamiltonian \(H\) [1303.0578]. The corresponding Heisenberg dynamics satisfy
\[
dj_t(X)=j_t(\mathcal{L}_{11}X)\,d\Lambda(t) +j_t(\mathcal{L}_{10}X)\,dB^*(t) +j_t(\mathcal{L}_{01}X)\,dB(t) +j_t(\mathcal{L}_{00}X)\,dt,
\]
with Evans–Hudson maps \(\mathcal L_{jk}\) and Lindblad generator
\[
\mathcal{L}_{00}X=\frac12L^*[X,L]+\frac12[L^*,X]L-i[X,H].
\]
The input–output relations
\[
dB^{\mathrm{out}}(t)=j_t(S)\,dB(t)+j_t(L)\,dt,
\]
\[
d\Lambda^{\mathrm{out}}(t)=d\Lambda(t)+j_t(L^*S)\,dB^*(t)+j_t(S^*L)\,dB(t)+j_t(L^*L)\,dt
\]
provide the quantum analogue of classical state and observation equations [1303.0578].

Two derivational strategies recur throughout the literature. The first is the reference probability approach, the quantum analogue of the classical Girsanov/Kallianpur–Striebel method, based on the operator-valued relation
\[
\pi_t(X)=\frac{\sigma_t(X)}{\sigma_t(I)},
\]
where \(\sigma_t\) is an unnormalized conditional state [1303.0578]. The second is the characteristic-function or generating-map method, in which one introduces an adapted process \(C(t)\) or \(g(k,t)\), expands identities such as
\[
\mathbb{E}\big[(\pi_t(X)-j_t(X))C(t)\big]=0,
\]
and identifies the drift and stochastic gain by Itō calculus [1303.0578, 1212.3512].

A different but related algebraic viewpoint appears in Belavkin’s matrix representation of quantum stochastic calculus. There, QSDE coefficients are packaged into Belavkin matrices, the quantum Itō product becomes ordinary matrix multiplication, and physical realizability is encoded by the \(\star\)-unitarity condition
\[
\mathbb{V}\mathbb{V}^{\star}=\mathbb{I} = \mathbb{V}^{\star}\mathbb{V}.
\]
In that broader sense, “Belavkin equations” may also refer to structural matrix equations governing quantum stochastic evolutions and feedback networks [1005.5644].

## 3. Diffusive and counting Belavkin equations

In the Heisenberg-picture conditional-expectation form used for multichannel diffusive observations, the Belavkin–Kushner–Stratonovich equation takes the form
\[
d\pi_t(\xi) = \pi_t(\mathscr{L}(\xi))\,dt + \beta^T K\, d\chi,
\]
with innovation
\[
d\chi = dZ - 2FF^T K^T \pi_t(\operatorname{Re}M)\,dt,
\]
and
\[
\beta := \pi_t(M^\# \xi + \xi M) -2\pi_t(\xi)\pi_t(\operatorname{Re}M)
\]
[1602.07911]. This is the standard BKSE for conditional expectations, and phase-space work interprets it as the starting point for posterior quasi-characteristic-function dynamics [1602.07911].

For repeated indirect measurements, the continuous diffusive limit yields the normalized stochastic master equation
\[
d\rho_t= L_d(\rho_t)\, dt + \sum_j \mathcal{D}_j(\rho_t)\, dX_{t}(j),
\]
with
\[
L_d(\rho_t):=-i[H_{s},\rho_t]+\sum_{i}p_{0}(i)\Big(C_{i}\,\rho_t\, C_{i}^{\dagger}-\frac{1}{2}\{C_{i}^{\dagger}C_{i},\rho_t\}\Big),
\]
and
\[
\mathcal{D}_j(\rho_t):=C_{j}\rho_{t}+\rho_{t}C_{j}^{\dagger}-\rho_{t}\, {\rm Tr}[(C_{j}+C_{j}^{\dagger})\rho_{t}],
\]
while the jump limit yields
\[
d\rho_t= L_p(\rho_t)\, dt +\sum_{i\not=i^*} \widehat{\mathcal{D}_i(\rho_t)}\, dY_{t}(i)
\]
with
\[
\widehat{\mathcal{D}_i(\rho_t)}= \frac{D_{i}\,\rho_{t}\,D_{i}^{\dagger}} {{\rm Tr}[D_{i}\rho_{t}D_{i}^{\dagger}]} -\rho_t
\]
[1210.0425]. This makes explicit that diffusive and counting Belavkin equations are scaling limits of discrete state updates.

Concrete state-vector and density-matrix forms appear in oscillator models. For intense balanced heterodyne detection of a single harmonic oscillator coupled to a coherent input field \(e(f)\), the unnormalized posterior wave function obeys
\[
\mathrm d\widehat\psi(t) = -\left( K+\sqrt{\mu}\bigl(a^\dagger f(t)-a\,\overline{f(t)}\bigr) \right)\widehat\psi(t)\,\mathrm dt + \sqrt{\mu}\,a\,e^{-i\phi(t)}\widehat\psi(t)\,\mathrm dW(t),
\]
with
\[
\mathrm dW(t)=\mathrm d\mathcal Q(t)-2\,\mathrm{Re}\!\left(e^{-i\phi(t)}f(t)\right)\mathrm dt,
\]
and the normalized posterior density matrix satisfies the corresponding nonlinear diffusive Belavkin equation [1212.3512].

## 4. Coherent, heterodyne, and squeezed-field generalizations

A central extension of the vacuum-input theory replaces the input bosonic field by a coherent state with amplitude \(\beta(t)\). In that case the effective coupling is shifted to
\[
L^{\beta(t)}=S\beta(t)+L,
\]
and the unconditional generator becomes
\[
\mathcal{L}^{\beta(t)}X = \mathcal{L}_{00}X+\beta(t)^*\mathcal{L}_{10}X+\beta(t)\mathcal{L}_{01}X+|\beta(t)|^2\mathcal{L}_{11}X.
\]
The resulting filters for both quadrature measurements and photon counting have the same structural form as the vacuum Belavkin equations, but with \(L\) replaced by \(L^{\beta(t)}\) and \(\mathcal L_{(L,H)}\) replaced by \(\mathcal L^{\beta(t)}\) [1303.0578].

For quadrature measurement, the normalized coherent-state filter is
\[
d\pi_t (X) = \pi_t \big(\mathcal{L}^{\beta(t)}X\big) \, dt + \Big\{\pi_t \big(XL^{\beta(t)} + L^{\beta(t)*}X\big) - \pi_t (X) \pi_t \big(L^{\beta(t)} + L^{\beta(t)*}\big) \Big\} \, dI^{\mathrm{quad}}(t),
\]
with innovations
\[
dI^{\mathrm{quad}} (t) = dY^{\mathrm{out}} (t) - \pi_t \big(L^{\beta(t)} + L^{\beta(t)*}\big) dt.
\]
For photon counting,
\[
d\pi_t (X) = \pi_t \big(\mathcal{L}^{\beta(t)} X\big) dt + \left\{ \frac{\pi_t \big(L^{\beta(t)*} X L^{\beta(t)}\big)} {\pi_t \big(L^{\beta(t)*} L^{\beta(t)}\big)} - \pi_t (X) \right\} dI^{\mathrm{num}} (t),
\]
with
\[
dI^{\mathrm{num}} (t) = dY^{\mathrm{out}} (t) - \pi_t \big(L^{\beta(t)*}L^{\beta(t)}\big) dt.
\]
When \(\beta(t)=0\), these reduce exactly to the familiar vacuum Belavkin equations [1303.0578].

A heterodyne realization for a coherent input field shows additional dynamical consequences. In the cavity-mode model of continuous diffusion observation, the filtering equation is relaxing: any initial square-integrable wave function tends asymptotically to a coherent state whose amplitude depends on the coupling constant \(\mu\) and the initial coherent state \(e(f)\) of the apparatus through \(f(t)\) [1212.3512]. For squeezed coherent initial states, the posterior state remains within the squeezed coherent family, the squeezing parameter satisfies a Riccati equation, and asymptotically \(\Gamma(t)\to0\), so the state becomes coherent [1212.3512].

Belavkin filtering also extends to squeezed Gaussian fields. When monitored outputs are mixed with squeezed light, or when the system is driven directly by squeezed noise, the innovation channels acquire a nontrivial covariance
\[
dY_\alpha\, dY_\beta = K_{\alpha\beta}\,dt,
\]
and the gain is weighted by \(K^{-1}\). The generalized filter is
\[
d\pi _{t}(X)=\pi _{t}(\mathcal{L}X)\,dt+\mathcal{H}_{t}^{\alpha }\left( X\right) dI_{\alpha }\left( t\right),
\]
where \(\mathcal H_t^\alpha(X)\) contains both the usual homodyne covariance terms and additional commutator terms involving squeezed-state correlations \(N\) and \(M\) [1405.7795]. This shows that squeezing alters not only the observation covariance but also the unconditional generator itself.

## 5. Alternative formulations, repeated-measurement limits, and structural generalizations

One major reformulation replaces operator-valued filtering dynamics by phase-space evolution. For nonlinear quantum stochastic systems with Weyl-quantized Hamiltonian and coupling operators, the posterior quasi-characteristic function
\[
\Phi(t,u):=\pi_t(W_u)
\]
satisfies the stochastic integro-differential equation
\[
d\Phi(t,u) = A(\Phi(t,\cdot))(u)\,dt + B(\Phi(t,\cdot))(u)^T K\, d\chi.
\]
This is a spatial Fourier-domain representation of the Belavkin–Kushner–Stratonovich equation, and its inverse Fourier transform yields a corresponding equation for the posterior quasi-probability density [1602.07911]. In the linear-Gaussian case, the same framework reproduces the quantum Kalman filter [1602.07911].

A second formulation starts from repeated indirect measurements. In that picture, a complete measurement consists of infinitely many partial measurements, each partial outcome updates the state by Bayes’ rule, and the continuous-time Brownian or Poisson limits produce the diffusive and jump Belavkin equations [1210.0425]. This discrete-to-continuous derivation makes the innovation process directly interpretable as the centered measurement record. A more recent repeated-measurement analysis shows that the same construction can also converge to non-Markovian Volterra stochastic equations when memory is introduced at the microscopic level; in that case the limit is no longer a standard Markovian Belavkin SDE but a Volterra-type stochastic equation with memory kernel [2512.11462].

A third structural extension uses Belavkin matrices. In the Gough–James theory of quantum feedback networks, the coefficient matrix
\[
\mathbb{V}= \begin{pmatrix} 1 & -L^\dagger S & -\frac12 L^\dagger L-iH\\ 0 & S & L\\ 0 & 0 & 1 \end{pmatrix}
\]
encodes the Hudson–Parthasarathy coefficients \((S,L,H)\), and physical realizability is equivalent to
\[
\mathbb{V}\mathbb{V}^{\star}=\mathbb{I} = \mathbb{V}^{\star}\mathbb{V}.
\]
Feedback reduction then becomes a non-commutative Möbius transformation
\[
\mathcal{F}(\mathbb{V},X)_{\alpha\beta} = V_{\alpha\beta} + V_{\alpha i}\,X\,(1-V_{ii}X)^{-1}\,V_{i\beta},
\]
which again preserves \(\star\)-unitarity [1005.5644]. In this broader algebraic usage, Belavkin equations refer to matrix identities governing quantum stochastic evolutions and feedback interconnections.

## 6. Many-body, mean-field, infinite-dimensional, and terminological developments

Belavkin equations have also been extended to many-particle systems under continuous observation. For \(N\) identical particles with mean-field interaction, the normalized \(N\)-particle state obeys a many-body stochastic Schrödinger equation or, equivalently, a many-body stochastic master equation, and as \(N\to\infty\) the one-particle marginal converges to a nonlinear stochastic equation of McKean–Vlasov type [2008.07375]. In density-matrix form, the limiting equation is
\[
d\gamma_{j,t} = -i[H+u(t,\gamma_{j,t})\hat H,\gamma_{j,t}]\,dt -i[A^{\bar\eta_t},\gamma_{j,t}]\,dt + \left( L\gamma_{j,t}L^* -\frac12L^*L\gamma_{j,t} -\frac12\gamma_{j,t}L^*L \right)dt
\]
\[
\qquad + \Bigl( \gamma_{j,t}L^* +L\gamma_{j,t} -\gamma_{j,t}\operatorname{tr}(\gamma_{j,t}(L+L^*)) \Bigr)dB_t,
\]
with \(\eta_t=\mathbb E\,\gamma_{j,t}\) [2008.07375]. This gives a quantum mean-field analogue of a McKean–Vlasov diffusion.

A finite-dimensional controlled version writes the mean-field Belavkin equation directly for density matrices,
\[
\mathrm{d}\gamma_t= (-i[ H + u(\gamma_t)\hat{H} + A^{m}_t, \gamma_t] )\mathrm{d}t + \left(L\gamma_tL^{\dag} - \frac{1}{2} \{L^{\dag}L,\gamma_t\}\right)\mathrm{d}t
\]
\[
\qquad +\sqrt{\eta}\Big(\gamma_tL^{\dag} + L\gamma_t - \mathrm{tr}\big((L + L^{\dag})\gamma_t\big)\gamma_t\Big)\mathrm{d}W_t,
\]
with \(m_t=\mathbb E[\gamma_t]\) [2303.09667]. Under boundedness and Lipschitz assumptions, the equation is well posed and, for \(\eta=1\), propagation of chaos is proved under purification assumption [2303.09667].

The infinite-dimensional wave-function version on \(L^2(\mathbb R^d)\) has now been derived rigorously. For \(N\) interacting particles under continuous diffusive measurement, the mean-field limit is the stochastic Hartree-type Belavkin equation
\[
du(t) = \left( -i\big(H + V * \mathbb E|u(t)|^2\big) -\frac12\big(L^*L - 2\langle L\rangle_{u(t)}L + \langle L\rangle_{u(t)}^2\big) \right)u(t)\,dt
 + \big(L-\langle L\rangle_{u(t)}\big)u(t)\,d\beta_t,
\]
with global well-posedness proved directly by fixed-point methods and convergence from the \(N\)-body stochastic Schrödinger dynamics established in trace norm for reduced marginals [2507.19231]. A related mixed-state construction with unbounded Hamiltonians and unbounded interaction operators uses a system of stochastic interacting wave functions and proves that the reconstructed mixed state satisfies the diffusive stochastic quantum master equation, which is also known as Belavkin equation [2503.15280].

A recurrent source of confusion is terminological. Several recent papers concern the Belavkin–Staszewski relative entropy, BS-conditional mutual information, or BS-quantum Markov chains; these works are about quantum information divergences and recoverability, not about Belavkin stochastic filtering equations [1904.10768, 2501.09708]. This suggests a useful distinction between Belavkin equations in the filtering sense—stochastic evolution equations for conditioned quantum states—and Belavkin–Staszewski constructions in the information-theoretic sense.

Source: https://www.emergentmind.com/topics/belavkin-equations