---
title: Operator-Valued Backward Stochastic Integral Equations
url: https://www.emergentmind.com/topics/operator-valued-backward-stochastic-integral-equations
type: topic
---

# Operator-Valued Backward Stochastic Integral Equations

Operator-valued backward stochastic integral equations are backward stochastic equations in which the unknowns, coefficients, or martingale integrands act as operators on a Hilbert space rather than as scalar or vector processes. In infinite dimensions, they typically appear as operator-valued backward stochastic evolution equations, backward stochastic Riccati equations, or related mild backward integral systems with terminal condition at time \(T\). The subject combines three distinct but tightly connected ingredients: a stochastic integration theory for operator-valued integrands, solution concepts that remain meaningful when the natural operator space is not Hilbertian, and semigroup-based well-posedness mechanisms for concrete infinite-dimensional equations [1204.3275] [1404.6543] [1605.06549].

## 1. Formal structure and mathematical meaning

A canonical operator-valued backward stochastic evolution equation on a Hilbert space \(H\) has the formal form
\[
\begin{cases}
dP = - (A^*+J^*)P\,dt - P(A+J)\,dt - K^*PK\,dt -(K^*Q+QK)\,dt +F\,dt +Q\,dw(t), & \text{in } [0,T),\\
P(T)=P_T,
\end{cases}
\]
where \(P(t)\in \mathcal L(H)\) is the operator-valued unknown and \(Q\) is the martingale coefficient [1204.3275]. In control-theoretic applications, this equation is the infinite-dimensional analogue of the finite-dimensional matrix-valued second-order adjoint BSDE.

A more specialized but structurally central class is the operator-valued backward stochastic Riccati equation
\[
-dP(t) = \Big( A^*P(t)+P(t)A +C^*(t)P(t)C(t) +C^*(t)Q(t)+Q(t)C(t) -P(t)B(t)B^*(t)P(t) +S(t) \Big)\,dt - Q(t)\,dW(t),
\]
with terminal condition
\[
P(T)=M,
\]
posed on an infinite-dimensional Hilbert space and interpreted through a mild backward integral formula [1404.6543].

The phrase “backward stochastic integral equation” emphasizes that these objects are not merely differential expressions written backward in time. Their rigorous content is usually given either by a mild variation-of-constants representation,
\[
P(t)=\text{terminal term}+\int_t^T (\cdots)\,ds+\int_t^T (\cdots)\,dW(s),
\]
or by a duality identity against forward test equations [1204.3275] [1404.6543]. In this sense, the integral formulation is primary and the differential notation is often shorthand.

Two recurring features distinguish the operator-valued setting from standard Hilbert-valued BSDEs. First, the unknown \(P(t)\) acts on the state space \(H\), so the drift contains noncommutative compositions such as \(A^*P+PA\), \(C^*PC\), \(C^*Q+QC\), and \(PBB^*P\). Second, in infinite dimensions the natural space \(L(H)\) of bounded operators is a Banach space that is not, in general, an adequate Hilbertian environment for classical stochastic calculus [1204.3275] [1404.6543].

## 2. Operator-valued stochastic integration as a foundational layer

A foundational integration framework is provided by the Hilbert-space-valued stochastic integral of operator-valued functions introduced in "A stochastic integral of operator-valued functions" [1605.06549]. The setting starts with a complex Hilbert space \(\mathcal H\), a nonzero vector \(M\in\mathcal H\), and a right-continuous increasing family of orthogonal projections
\[
[0,T]\ni t\mapsto E_t\in \mathcal L(\mathcal H),
\qquad E_T=1,
\]
which determines a projector-valued measure \(E(\cdot)\). The associated abstract martingale is
\[
M_t:=E_tM,\qquad t\in[0,T],
\]
and the corresponding \(\mathcal H\)-valued set function is
\[
M(\alpha):=E(\alpha)M.
\]
The control measure is
\[
\mu(\alpha):=\|M(\alpha)\|_{\mathcal H}^2 =(E(\alpha)M,M)_{\mathcal H},\qquad \alpha\in\mathcal B([0,T]).
\]

The integrands are operator-valued. For \(t\in[0,T]\), the increment subspace is
\[
\mathcal H_M(t):=\operatorname{span}\{M_{s_2}-M_{s_1}\mid (s_1,s_2]\subset (t,T]\}\subset \mathcal H,
\]
and the relevant operator space is
\[
\mathcal L_M(t):=\mathcal L(\mathcal H_M(t)\to \mathcal H).
\]
An operator is \(\mathcal L_M(t)\)-measurable if it is bounded on \(\mathcal H_M(t)\), its operator norm is consistent across later times, and it partially commutes with the resolution of identity:
\[
AE_sg=E_sAg,\qquad g\in\mathcal H_M(t),\ s\in[t,T].
\]
This is the paper’s substitute for adaptedness.

For a simple \(\mathcal L_M\)-adapted operator-valued function
\[
A(t)=\sum_{k=0}^{n-1}A_k\,\varkappa_{(t_k,t_{k+1}]}(t),
\]
the stochastic integral is defined by
\[
\int_{[0,T]}A(t)\,dM_t :=\sum_{k=0}^{n-1}A_k\bigl(M_{t_{k+1}}-M_{t_k}\bigr).
\]
The \(L^2\)-type quasinorm on simple integrands is
\[
\|A\|_{S_2} :=\left(\int_{[0,T]}\|A(t)\|_{\mathcal L_M(t)}^2\,d\mu(t)\right)^{1/2},
\]
and completion yields the admissible class \(S_2=S_2(M)\). The key estimate is
\[
\left\|\int_{[0,T]}A(t)\,dM_t\right\|_{\mathcal H}^2 \le \int_{[0,T]}\|A(t)\|_{\mathcal L_M(t)}^2\,d\mu(t),
\]
together with linearity of the integral. Orthogonality of projector increments and the partial commutation condition are the structural devices behind this bound.

This construction is not itself a theory of backward stochastic equations. The paper explicitly does not formulate or solve BSDEs or backward stochastic integral equations. Nevertheless, it shows that the abstract integral recovers both the classical Itô integral with respect to a normal martingale and the Itô integral in symmetric Fock space. In the normal-martingale realization, \(\mathcal L_M(t)\)-measurability coincides exactly with ordinary adaptedness of multiplication operators, and the admissible class \(S_2\) becomes the classical \(L_a^2\) space [1605.06549]. This suggests that operator-valued backward equations can be supplied with a rigorous stochastic integral term once an appropriate notion of operator adaptedness and \(L^2\)-control is available.

## 3. Transposition and relaxed transposition formulations

A general theory for operator-valued backward stochastic evolution equations is developed by Lü and Zhang through transposition methods [1204.3275]. The central motivation is that direct \(\mathcal L(H)\)-valued stochastic integration is unavailable in the needed level of generality. The paper works on a complete filtered probability space with a one-dimensional standard Brownian motion and explicitly allows a general filtration, without assuming natural filtration or quasi-left continuity.

The key obstacle is functional-analytic. In the infinite-dimensional setting, \(\mathcal L(H)\) is nonreflexive and nonseparable, so one cannot directly invoke the standard Hilbertian martingale representation or the usual BSDE machinery. To bypass this, the equation is defined through bilinear identities against forward stochastic evolution equations. For forward test systems
\[
\begin{cases}
dx_1=(A+J)x_1\,ds + u_1\,ds + Kx_1\,dw(s)+v_1\,dw(s),\\
x_1(t)=\xi_1,
\end{cases}
\qquad
\begin{cases}
dx_2=(A+J)x_2\,ds + u_2\,ds + Kx_2\,dw(s)+v_2\,dw(s),\\
x_2(t)=\xi_2,
\end{cases}
\]
the operator-valued transposition solution \((P,Q)\) is characterized by the identity
\[
\begin{aligned}
&\mathbb E\langle P_Tx_1(T),x_2(T)\rangle_H -\mathbb E\int_t^T \langle F(s)x_1(s),x_2(s)\rangle_H\,ds \\
={}& \mathbb E\langle P(t)\xi_1,\xi_2\rangle_H +\mathbb E\int_t^T \langle P(s)u_1(s),x_2(s)\rangle_H\,ds +\mathbb E\int_t^T \langle P(s)x_1(s),u_2(s)\rangle_H\,ds \\
&+\mathbb E\int_t^T \langle P(s)K(s)x_1(s),v_2(s)\rangle_H\,ds +\mathbb E\int_t^T \langle P(s)v_1(s),K(s)x_2(s)+v_2(s)\rangle_H\,ds \\
&+\mathbb E\int_t^T \langle Q(s)v_1(s),x_2(s)\rangle_H\,ds +\mathbb E\int_t^T \langle Q(s)x_1(s),v_2(s)\rangle_H\,ds .
\end{aligned}
\]
The backward martingale term is therefore encoded by duality rather than by constructing a strong \(\mathcal L(H)\)-valued stochastic integral.

The theory has two levels. In the Hilbert–Schmidt setting, where
\[
P_T\in L^2(\Omega;\mathcal L_2(H)),\qquad F\in L^1(0,T;L^2(\Omega;\mathcal L_2(H))),
\]
the equation admits a unique transposition solution
\[
(P,Q)\in D_{\mathbb F}([0,T];L^2(\Omega;\mathcal L_2(H)))\times L^2_{\mathbb F}(0,T;\mathcal L_2(H))
\]
and the problem can be treated by lifting it to the Hilbert space \(\mathcal L_2(H)\) [1204.3275].

For general bounded-operator data, the paper proves uniqueness of transposition solutions but establishes existence only in a relaxed sense. The relaxed transposition solution replaces a bona fide process \(Q(\cdot)\in \mathcal L(H)\) by operator families
\[
(Q^{(\cdot)},\widehat Q^{(\cdot)})\in \mathcal Q[0,T],
\]
which act on forward test triples and enter a modified duality identity. Theorem 6.1 yields a unique relaxed transposition solution
\[
(P(\cdot),Q^{(\cdot)},\widehat Q^{(\cdot)}) \in D_{\mathbb F,w}([0,T];L^2(\Omega;\mathcal L(H)))\times \mathcal Q[0,T].
\]
This is one of the defining features of the subject: in full operator generality, the martingale integrand may be representable only indirectly.

The methodological core consists of finite-dimensional approximations, Yosida regularization, and new weakly sequential Banach-Alaoglu-type compactness results for uniformly bounded operator families. Those compactness theorems make it possible to pass from matrix-valued approximating BSDEs to an infinite-dimensional operator-valued limit [1204.3275].

## 4. Mild semigroup formulations and operator-valued BSREs

A different resolution of the operator-valued backward problem is given by the mild theory of infinite-dimensional backward stochastic Riccati equations [1404.6543]. The underlying space is a real separable Hilbert space \(H\), and the unbounded drift operator \(A\) is assumed self-adjoint, diagonalizable, and such that
\[
Ae_k=-\lambda_k e_k,\qquad \omega\le \lambda_1\le \lambda_2\le \cdots,
\]
with
\[
\sum_{k\ge1}\lambda_k^{-2\rho}<\infty
\qquad\text{for some }\rho\in\Big(\frac14,\frac12\Big).
\]
Under these assumptions, \(A\) generates an analytic contraction semigroup \((e^{tA})_{t\ge0}\).

The decisive idea is not to seek the martingale coefficient \(Q\) in \(L(H)\) itself. Instead, the paper introduces
\[
V=D((-A)^\rho),\qquad V'=D((-A)^{-\rho}),
\]
forming a Hilbert triple
\[
V \hookrightarrow H \hookrightarrow V',
\]
and defines the Hilbert space of operators
\[
\mathcal K := L_2(V;H)\cap L_2(H;V'),
\qquad
|T|_{\mathcal K}^2 := |T|_{L_2(V;H)}^2 + |T|_{L_2(H;V')}^2.
\]
A central structural fact is
\[
L(H)\subset \mathcal K.
\]
Its symmetric subspace \(\mathcal K_s\) is used for \(Q\). The semigroup regularization bounds
\[
t^\rho |e^{tA}|_{L(H,V)}\le 1,\qquad t^\rho |e^{tA}|_{L(V',H)}\le M_A
\]
are then used to control sandwiched operator expressions of the form
\[
e^{(s-t)A^*}(C^*Q+QC)e^{(s-t)A},
\]
even though \(QC\) need not belong to \(\mathcal K\).

For the Lyapunov equation, the mild solution is a pair
\[
(P,Q)\in L^2_{\mathcal P,S}\big(\Omega;C([0,T];\Sigma(H))\big)\times L^2_{\mathcal P}(\Omega\times[0,T];\mathcal K_s)
\]
satisfying
\[
\begin{aligned}
P(t) &= e^{(T-t)A^*} M e^{(T-t)A}
 + \int_t^T e^{(s-t)A^*} S(s)e^{(s-t)A}\,ds \\
&\quad + \int_t^T e^{(s-t)A^*} \Big[ C^*(s)P(s)C(s)+C^*(s)Q(s)+Q(s)C(s) \Big]e^{(s-t)A}\,ds \\
&\quad + \int_t^T e^{(s-t)A^*}Q(s)e^{(s-t)A}\,dW(s).
\end{aligned}
\]
Theorem 3.4 proves existence and uniqueness of this mild solution.

For the Riccati equation, the mild form becomes
\[
\begin{aligned}
P(t) &= e^{(T-t)A^*} M e^{(T-t)A}
 + \int_t^T e^{(s-t)A^*} S(s)e^{(s-t)A}\,ds \\
&\quad + \int_t^T e^{(s-t)A^*} \Big[ C^*(s)P(s)C(s) -P(s)B(s)B^*(s)P(s) +C^*(s)Q(s)+Q(s)C(s) \Big] e^{(s-t)A}\,ds \\
&\quad + \int_t^T e^{(s-t)A^*}Q(s)e^{(s-t)A}\,dW(s),
\qquad \mathbb P\text{-a.s.}
\end{aligned}
\]
Theorem 4.4 establishes existence and uniqueness of a mild solution \((P,Q)\) on \([0,T]\), with
\[
P\in L^\infty_{\mathcal P,S}(\Omega\times(0,T);\Sigma^+(H)).
\]

The proof is local-to-global. Regularized approximants based on
\[
J_n = n(nI-A)^{-1}
\]
produce Hilbert-Schmidt-valued backward equations, for which standard Hilbert-space BSDE arguments are available. A priori estimates are obtained first on short backward intervals, including the local bound
\[
\|P\|_{L^2(\Omega;C([T-\delta,T];L(H)))}^2 + \mathbb E\int_{T-\delta}^T |Q(s)|_{\mathcal K}^2\,ds
\le c\Big( \mathbb E|M|_{L(H)}^2 + \mathbb E\int_{T-\delta}^T |S(s)|_{L(H)}^2\,ds \Big),
\]
and continuation then yields global existence [1404.6543].

## 5. Control-theoretic roles: second-order adjoints and feedback operators

Operator-valued backward equations enter stochastic control in two distinct but related ways.

In the Pontryagin framework for general stochastic evolution equations,
\[
\begin{cases}
dx = [Ax + a(t,x,u)]\,dt + b(t,x,u)\,dw(t),\\
x(0)=x_0,
\end{cases}
\]
with cost
\[
J(u(\cdot)) \triangleq \mathbb E\left[\int_0^T g(t,x(t),u(t))\,dt+h(x(T))\right],
\]
the first adjoint is vector-valued, but control-dependent diffusion and nonconvex control domains require a second-order adjoint equation that is operator-valued [1204.3275]. The Hamiltonian is
\[
\mathbb H(t,x,u,k_1,k_2) = \langle k_1,a(t,x,u)\rangle_H + \langle k_2,b(t,x,u)\rangle_H - g(t,x,u),
\]
and the second-order adjoint data are
\[
P_T=-h_{xx}(\bar x(T)),\qquad
J(t)=a_x(t,\bar x(t),\bar u(t)),\qquad
K(t)=b_x(t,\bar x(t),\bar u(t)),
\]
\[
F(t)=-\mathbb H_{xx}(t,\bar x(t),\bar u(t),y(t),Y(t)).
\]
The resulting operator-valued BSEE supplies the quadratic correction term in the maximum condition,
\[
\operatorname{Re}\,\mathbb H(t,\bar x(t),\bar u(t),y(t),Y(t))
-\operatorname{Re}\,\mathbb H(t,\bar x(t),u,y(t),Y(t))
-\frac12 \langle P(t)\delta b(t),\delta b(t)\rangle_H \ge 0,
\]
where
\[
\delta b(t)=b(t,\bar x(t),\bar u(t))-b(t,\bar x(t),u).
\]
This term records the second-order effect of diffusion perturbations and is indispensable in the nonconvex, control-dependent diffusion regime.

In stochastic linear-quadratic control, the operator-valued BSRE plays the synthesis role usually occupied by the Riccati equation in deterministic control. The controlled state equation is
\[
dy(t)=(Ay(t)+B(t)u(t))\,dt + C(t)y(t)\,dW(t), \qquad y(0)=x,
\]
with cost
\[
J(u)=\mathbb E\int_0^T \big( \langle S(s)y(s),y(s)\rangle_H + |u(s)|_U^2 \big)\,ds + \mathbb E\,\langle M y(T),y(T)\rangle_H.
\]
The mild BSRE solution \(P\) determines the value function through
\[
J(0,x,\bar u)=\langle P(0)x,x\rangle_H,
\]
and the optimal feedback law is
\[
\bar u(s)=-B^*(s)P(s)\bar y(s).
\]
The closed-loop state equation is then
\[
d\bar y(s)=\big(A-B(s)B^*(s)P(s)\big)\bar y(s)\,ds + C(s)\bar y(s)\,dW(s),\qquad \bar y(0)=x.
\]
Thus, in the maximum-principle setting the backward operator equation functions as a second-order adjoint, whereas in the LQ setting it functions as a feedback-generating Riccati equation [1204.3275] [1404.6543].

## 6. Functional-analytic difficulties, scope, and limitations

The theory is shaped by several obstructions that do not arise in finite-dimensional matrix-valued BSDEs.

The first is that the natural operator space \(L(H)\) is only a Banach space and, in infinite dimensions, is nonseparable and nonreflexive. Lü and Zhang explicitly identify this as the reason classical operator-valued martingale representation and standard BSDE methods fail in the general bounded-operator setting [1204.3275]. Their response is the transposition framework and, in full generality, the relaxed transposition solution. A common misconception is that every operator-valued backward stochastic equation should admit a strong \(Q(t)\in \mathcal L(H)\)-valued martingale integrand; the general theory in fact does not provide this. It provides a surrogate object \( (Q^{(\cdot)},\widehat Q^{(\cdot)}) \) defined through dual operator families.

The second is that even when one enlarges the operator space to a Hilbert space, algebraic drift terms need not preserve that space. In the BSRE theory, \(Q\) is placed in
\[
\mathcal K=L_2(V;H)\cap L_2(H;V'),
\]
but terms such as \(QC\) need not belong to \(\mathcal K\). The solution is not to close the drift in \(\mathcal K\) abstractly, but to exploit analytic semigroup smoothing after sandwiching by \(e^{(s-t)A}\) and \(e^{(s-t)A^*}\) [1404.6543]. This is a highly specific infinite-dimensional mechanism rather than a generic feature of operator-valued BSDEs.

The third is that stochastic integration of operator-valued integrands requires an adaptedness notion that remains meaningful beyond scalar or vector processes. Tesko’s construction supplies exactly such a notion through \(\mathcal L_M(t)\)-measurability and an \(L^2\)-type control measure, but the paper does not develop backward equations or existence and uniqueness for operator-valued BSDEs [1605.06549]. It is therefore best interpreted as a preparatory integration theory.

These three strands delineate the current scope reflected in the cited works. One strand gives an abstract operator-valued stochastic integral with classical and Fock-space specializations. A second gives a general operator-valued backward evolution theory in duality form, with relaxed existence in the bounded-operator regime. A third gives direct mild well-posedness for Lyapunov and Riccati equations under analytic semigroup regularization. This suggests that a fully unified theory of operator-valued backward stochastic integral equations would need to combine operator-valued integration, weak or mild notions of solution, and problem-specific regularization mechanisms; however, that synthesis is not proved in these papers [1204.3275] [1404.6543] [1605.06549].

Source: https://www.emergentmind.com/topics/operator-valued-backward-stochastic-integral-equations