---
title: 'Non-Markovian MFGs: Memory and Dynamics'
url: https://www.emergentmind.com/topics/non-markovian-mean-field-games
type: topic
---

# Non-Markovian MFGs: Memory and Dynamics

Non-Markovian mean-field games are mean-field game models in which the representative agent’s optimization problem or the consistency condition depends on temporal history rather than only on the current state. In the supplied literature, non-Markovianity appears in several distinct forms: subdiffusive dynamics generated by an inverse stable subordinator and encoded by Caputo time-fractional derivatives [1812.05431]; path-dependent interactions through the empirical sub-probability measure of survivors and the history of absorptions [1902.02670]; weak-form games with coefficients depending on the joint distribution of states and controls on path space [2105.00484]; and recursive utility portfolio games with Epstein-Zin preferences formulated without any Markov structure and characterized by BSDEs [2505.07231]. Across these formulations, the common objective is to define Nash equilibria in a continuum of strategically interacting agents when memory, delayed effects, absorption history, or non-separable intertemporal preferences invalidate the classical dynamic programming paradigm.

## 1. Core formulations of non-Markovianity

A first class of non-Markovian mean-field games arises from anomalous diffusion. In the time-fractional formulation, the microscopic state process is obtained by time-changing a diffusion \(Y_s\) with the inverse \(E_t\) of a strictly increasing \(\beta\in(0,1)\)-stable subordinator \(D_t\). The resulting trajectory
\[
X_t=Y_{E_t}, \qquad dY_s=v(s,Y_s)\,ds+\sqrt{2}\,dB_s
\]
is continuous, non-Markovian, non-Gaussian, and carries a long-term memory because \(E_t\) “pauses” the motion for random heavy-tailed times [1812.05431]. In this setting, the law \(m(t,x)\) satisfies a time-fractional Fokker-Planck equation, and equilibrium is described by a coupled time-fractional HJB-FP system.

A second class is path dependence through absorption. In the framework of games with smooth dependence on past absorptions, each player’s trajectory lives in \(X=C([0,T];\mathbb{R}^d)\), with absorption time
\[
\tau^\varphi=\inf\{t\in[0,T]:\varphi(t)\notin O\}\wedge T,
\]
and interaction enters through the non-normalized empirical sub-probability measure of survivors,
\[
\mu_t^N=\frac1N\sum_{i=1}^N \delta_{X_t^{N,i}}\,1_{\{t<\tau^{N,i}\}},
\]
together with the fraction absorbed,
\[
A_t^N=1-\mu_t^N(\mathbb{R}^d).
\]
In the MFG limit, both \(\mu_t\) and \(A_t=\int_0^t(1-\mu_s(\mathbb{R}^d))\,ds\) retain the history of losses from the game [1902.02670].

A third class is fully non-Markovian weak-form control. Here the state process is specified on path space \(C_m=C([0,T];\mathbb{R}^m)\), the control is an \(A\)-valued predictable process, and the controlled law is defined by Girsanov transformation. The coefficients may depend on the full stopped path \(X_{\cdot\wedge s}\) and on the time-indexed law flow \(\xi_s\in P_2(C_m\times A)\), allowing interaction through the joint distribution of players’ states and controls [2105.00484].

A fourth class comes from recursive utility. In mean field portfolio games with Epstein-Zin preferences, the game is developed “in a general non-Markovian framework” and “without assuming any Markov structure.” The representative agent’s utility is defined by a BSDE-type recursion with externality \(\nu\), and equilibrium requires
\[
\ln \nu_t^*=E[\ln C_t^* \mid \mathcal{F}_t^0], \qquad
\ln \nu_T^*=E[\ln X_T^* \mid \mathcal{F}_T^0].
\]
The non-Markovian feature is not only path dependence of coefficients, but also recursion in utility and conditioning on common noise [2505.07231].

These formulations show that “non-Markovian” is not a single modeling choice. It may refer to memory in the state dynamics, memory in the interaction term, a weak formulation on path space, or recursive intertemporal preferences.

## 2. Time-fractional mean-field games and subdiffusive memory

The time-fractional MFG system studied in “Variational time-fractional Mean Field Games” couples a backward time-fractional HJB equation for the value function \(u\) with a forward time-fractional Fokker-Planck equation for the density \(m\):
\[
-\,^C\!D_{T-}^\beta u(t,x)+H(x,\nabla u(t,x))-V(x,m(t,x))-\Delta u(t,x)=0,
\qquad u(T,x)=U_T(x),
\]
\[
^C\!D_{0+}^\beta m(t,x)-\nabla\cdot(m\,H_p(x,\nabla u))-\Delta m(t,x)=0,
\qquad m(0,x)=m_0(x).
\]
Here the forward and backward Caputo derivatives are nonlocal in time, since they are convolutions against the kernel \((t-s)^{-\beta}\), and therefore explicitly encode memory [1812.05431].

The forward Caputo derivative is
\[
^C\!D_{0+}^\beta\varphi(t)=\frac{1}{\Gamma(1-\beta)}\int_0^t (t-s)^{-\beta}\varphi'(s)\,ds,
\]
and the backward Caputo derivative on \([0,T]\) is
\[
^C\!D_{T-}^\beta\varphi(t)
=
-\frac{1}{\Gamma(1-\beta)}
\int_t^T (s-t)^{-\beta}\varphi'(s)\,ds.
\]
Because these operators are nonlocal in time, the instantaneous rate of change depends on the entire history. In the limit \(\beta\to1\), they recover the classical Markovian derivative, which the paper identifies with the kernel limit \(K_1(t)=\delta(t)\) and \(^C\!D^1m=m'\) [1812.05431].

In the quadratic case \(H(x,p)=\tfrac12|p|^2\), the system becomes
\[
-\,^C\!D_{T-}^\beta u+\tfrac12|\nabla u|^2-V(x,m)-\Delta u=0,
\qquad
^C\!D_{0+}^\beta m-\nabla\cdot(m\nabla u)-\Delta m=0.
\]
The paper interprets this as an extension of variational mean-field games to the subdiffusive situation, providing an Eulerian interpretation of time-fractional MFG systems [1812.05431].

Under the assumptions that \(f(x,m)\) is continuous in \((x,m)\), satisfies \(f(x,0)=0\), is increasing in \(m\), and the Lasry-Lions-type monotonicity condition
\[
\int (f(x,m_1)-f(x,m_2))(m_1-m_2)\ge 0
\]
holds, together with \(U_T\in C^2(\mathbb{T}^d)\) and \(m_0\in C^1>0\), the paper proves existence of a classical solution
\[
(u,m)\in C^{1,2}([0,T]\times\mathbb{T}^d)\times C^{1,2}([0,T]\times\mathbb{T}^d),
\]
uniqueness under monotonicity, and mass conservation and positivity: \(\int m(t)=1\) and \(m(t,x)>0\) for all \(t>0\) [1812.05431].

The motivating applications listed in the paper are anomalous transport in porous media and biological cells, latency and long-memory in high-frequency trading, and subdiffusive supply-demand dynamics. A plausible implication is that time-fractional MFGs are especially suited to strategic environments where waiting-time heterogeneity is itself part of the aggregate interaction.

## 3. Path-dependent interactions through absorption and survivor measures

The absorption framework develops a non-Markovian MFG in which players leave the game when their private states hit the boundary of a domain \(O\subset\mathbb{R}^d\). The controlled representative-player dynamics are
\[
X_t=X_0+\int_0^t b(s,X,\mu,u(s,X))\,ds+\sigma W_t,
\]
where \(b\) is obtained from
\[
\bar b:[0,T]\times\mathbb{R}^d\times\mathcal{M}_{\le1,1}(\mathbb{R}^d)\times\Gamma\to\mathbb{R}^d
\]
through the re-parameterization
\[
b(t,\varphi,\theta,u)=\bar b\bigl(t,\varphi(t),g(t,\theta),u\bigr),
\]
and
\[
\int_{\mathbb{R}^d}\psi(x)\,g(t,\theta)(dx)
=
\int_X \psi(\varphi(t))\,1_{[0,\tau^\varphi)}(t)\,\theta(d\varphi).
\]
This construction makes the interaction depend on the measure of surviving states and therefore on the cumulative history of absorptions [1902.02670].

The representative-player cost is
\[
J^\mu(u)=E\Bigl[\int_0^{\tau^X}\bar f(s,X_s,\mu_s,u(s,X))\,ds+F(\tau^X,X_{\tau^X})\Bigr].
\]
When \(\bar f(t,x,\mu,u)=\bar f_0(t,x,u)+\bar f_1(t,x,\mu)\), the paper states that one sees explicitly the dependence on the whole history \((\mu_s)_{s\le t}\) and equivalently on \(A_t\) [1902.02670].

A strict feedback MFG solution is a pair \((u,\mu)\) such that \(u\) minimizes \(J^\mu(\cdot)\) among all feedbacks and, if \(X\) solves the controlled SDE under \((u,\mu)\), then
\[
\mu_t(\cdot)=P\{X_t\in\cdot,\;\tau^X>t\}
\]
for all \(t\). The same formulation is extended to relaxed feedbacks \(\lambda:[0,T]\times X\to \mathcal{P}(\Gamma)\). In fixed-point form, the solution is characterized by
\[
u\in\arg\min_{v\in\mathcal{U}_{fb}}J^\mu(v),
\qquad
\mu_t(\cdot)=\mathrm{Law}(X_t;\,X\text{ under }(u,\mu))(\cdot\cap\{\tau^X>t\}),
\]
or equivalently through a law-operator \(\Psi\) whose fixed points are MFG solutions [1902.02670].

Under Lipschitz and sub-linear growth of \(\bar b,\bar f,F\) in \((x,\mu)\), continuity in the measure variable relative to \(1\)-Wasserstein at Wiener-absolutely-continuous laws, compactness of \(\Gamma\), nondegeneracy of \(\sigma\), admissible initial law \(\nu\), and a convexity assumption \((H_C)\) on the Hamiltonian minimizers, the paper proves existence of both a relaxed feedback solution \((\lambda,\mu)\) and, under an extra control-convexity assumption, a strict feedback solution \((u,\mu)\) [1902.02670].

Uniqueness is established under additional assumptions: \(\bar f\) splits as \(\bar f_0+\bar f_1\), the Lasry-Lions condition
\[
\int_{\mathbb{R}^d}
\bigl[\bar f_1(t,x,\mu)-\bar f_1(t,x,\tilde\mu)\bigr](\mu-\tilde\mu)(dx)\ge 0
\]
holds for all \(\mu,\tilde\mu\), the drift is independent of \(\mu\), and for each fixed \(\bar\mu\) the single-player control problem has a unique minimizer. In that case, any two feedback solutions coincide [1902.02670].

The same paper links the continuum model to the finite-\(N\) game. In the finite-dimensional interaction case
\[
\tilde b(t,x,\ell,m,u),\qquad \ell=1-\mu(\mathbb{R}^d),\quad m=\int w(x)\,\mu(dx),
\]
a feedback MFG solution induces an \(\varepsilon_N\)-Nash equilibrium for the \(N\)-player game with \(\varepsilon_N\to0\) as \(N\to\infty\). No explicit rate is given, although the text notes that in many examples one obtains \(\varepsilon_N=O(N^{-1/2})\) by standard law-of-large-numbers arguments [1902.02670].

## 4. Weak formulation, McKean-Vlasov BSDEs, and equilibrium characterization

The weak-form theory developed in “Non-asymptotic convergence rates for mean-field games: weak formulation and McKean-Vlasov BSDEs” considers a fully non-Markovian setting with drift control and interaction through the joint distribution of states and controls. Under a reference probability measure \(\mathbb{P}\), the uncontrolled coordinate process satisfies
\[
X_t=X_0+\int_0^t \sigma_s(X_{\cdot\wedge s})\,dW_s.
\]
Given a control \(\alpha\) and a path-law flow \(\xi=(\xi_t)_{t\le T}\in\mathfrak{P}\), the controlled measure \(\mathbb{P}^{\alpha,\xi}\) is defined by the Girsanov density
\[
\frac{d\mathbb{P}^{\alpha,\xi}}{d\mathbb{P}}
=
\exp\Bigl\{
\int_0^T b_s(X_{\cdot\wedge s},\xi_s,\alpha_s)\cdot dW_s
-\frac12\int_0^T |b_s(\cdots)|^2\,ds
\Bigr\},
\]
under which
\[
X_t
=
X_0+\int_0^t \sigma_s(X_{\cdot\wedge s})\,b_s(X_{\cdot\wedge s},\xi_s,\alpha_s)\,ds
+\int_0^t \sigma_s(X_{\cdot\wedge s})\,dW_s^{\alpha,\xi}.
\]
The drift is bounded, Lipschitz in path, measure, and control, and dissipative in the path variable:
\[
(x-x')\cdot\bigl(b_t(x,\xi,a)-b_t(x',\xi,a)\bigr)\le -K_b\|x-x'\|_\infty^2,
\qquad K_b>0.
\]
The reward data \(f\) and \(g\) are Borel, polynomial-growth, and Lipschitz in all arguments [2105.00484].

For a pair \((\hat\alpha,\xi)\), the objective is
\[
J^\xi(\alpha)
=
\mathbb{E}^{\mathbb{P}^{\alpha,\xi}}
\Bigl[
\int_0^T f_s(X_{\cdot\wedge s},\xi_s,\alpha_s)\,ds
+
g(X,\xi_T^1)
\Bigr],
\]
and a mean-field equilibrium is defined by optimality of \(\hat\alpha\) under \(\xi\) together with the consistency condition
\[
\xi_t=\mathrm{Law}^{\mathbb{P}^{\hat\alpha,\xi}}(X_{\cdot\wedge t},\hat\alpha_t),
\qquad \text{for a.e. } t.
\]
The paper then introduces the Hamiltonian
\[
h_t(x,\xi,z,a)=b_t(x,\xi,a)\cdot z+f_t(x,\xi,a),
\qquad
H_t(x,\xi,z)=\sup_{a\in A} h_t(x,\xi,z,a),
\]
with measurable selection \(\Lambda\in\arg\max_a h_t\), and proves that \((\hat\alpha,\xi)\) is a mean-field equilibrium if and only if there exist processes \((Y,Z)\) solving a generalized McKean-Vlasov BSDE, with
\[
\hat\alpha_t=\Lambda_t(X_{\cdot\wedge t},\,_{\hat\alpha}(X_{\cdot\wedge t},\hat\alpha_t),\,Z_t,\,0).
\]
Moreover \(Y_0=V^\xi\) [2105.00484].

This BSDE characterization is the central replacement for the Markovian HJB/master-equation route. Under the Lipschitz and dissipativity assumptions on \(b,f,\sigma\), and either a small terminal payoff \(\|g\|_\infty\le\Psi\ll1\) or a smooth terminal payoff in the measure argument, the BSDE admits a unique solution \((Y,Z)\) in
\[
\mathcal{S}^\infty(\mathbb{R})\times\mathcal{H}^2_{\mathrm{BMO}}.
\]
The proof combines change of measure, a fixed-point argument on \(\mathcal{S}^\infty\times\mathcal{H}^2_{\mathrm{BMO}}\times\{\text{measure-flows}\}\), contraction in a weighted norm, and the use of dissipativity to obtain global-in-\(T\) bounds [2105.00484].

A common misconception is that well-posedness and uniqueness in mean-field games necessarily require short time horizon, separability assumptions, or Lasry-Lions monotonicity. In this weak non-Markovian setting, the paper explicitly states that its existence and uniqueness results “do not require short time horizon, separability assumptions on the coefficients, nor Lasry and Lions's monotonicity conditions, but rather smallness, or alternatively regularity, conditions on the terminal reward and a dissipativity condition on the drift” [2105.00484].

## 5. Recursive utilities and non-Markovian portfolio games

The portfolio-game model with Epstein-Zin preferences provides a different non-Markovian mechanism, centered on recursive utility rather than on path-dependent state dynamics alone. On a filtered space carrying common noise \(W^0\) and idiosyncratic noise \(W^1,\dots,W^N\), the wealth dynamics of a representative agent are
\[
dX_t
=
r_tX_t\,dt+\Pi_t h_t\,dt+\Pi_t\sigma_t\,dW_t+\Pi_t\sigma_t^0\,dW_t^0-C_t\,dt,
\qquad X_0=x>0,
\]
with \(h_t=R_t-r_t\), bounded progressive coefficients, and \(\sigma^2+(\sigma^0)^2>0\) [2505.07231].

Given an externality \(\nu\), the Epstein-Zin utility satisfies the recursion
\[
V_t
=
E\Bigl[
\int_t^T f(C_s\nu_s^{-\theta},V_s)\,ds
+
\alpha U(X_T\nu_T^{-\theta})
\,\Bigm|\,\mathcal{F}_t
\Bigr],
\]
where
\[
U(x)=\frac{x^{1-\gamma}}{1-\gamma},
\]
\[
f(c,v)
=
\frac{\delta}{1-\frac1\psi}\,
c^{1-\frac1\psi}\,
\bigl((1-\gamma)v\bigr)^{1-\frac1{\widetilde\theta}}
-\delta\widetilde\theta\,v,
\]
\[
\widetilde\theta=\frac{1-\gamma}{1-1/\psi},
\qquad
\theta=\gamma-1.
\]
In the MFG limit, equilibrium requires the externality to coincide with the conditional log-average of optimal consumption and terminal wealth:
\[
\ln \nu_t^*=E[\ln C_t^*\mid\mathcal{F}_t^0],\qquad
\ln \nu_T^*=E[\ln X_T^*\mid\mathcal{F}_T^0].
\]
The paper establishes a uniqueness result by proving a one-to-one correspondence between Nash equilibria and the solutions to a class of BSDEs [2505.07231].

The core BSDE is
\[
-dY_t=g(t,Z_t,Z_t^0)\,dt-Z_t\,dW_t-Z_t^0\,dW_t^0,
\qquad
Y_T=-\theta(1-\gamma)\ln \nu_T,
\]
with the driver
\[
\begin{aligned}
g(t,z,z^0)
&=
\frac12\{z^2+(z^0)^2\}+(1-\gamma)r_t
+\frac{1-\gamma}{2\gamma}
\frac{(h_t+\sigma_t z+\sigma_t^0 z^0)^2}{\sigma_t^2+(\sigma_t^0)^2} \\
&\quad
+\frac{1-\gamma}{\psi-1}\,
\delta^\psi\,\nu_t^{-\theta(\psi-1)}\,(\alpha e^{Y_t})^{-\psi/\widetilde\theta}
-\delta\widetilde\theta.
\end{aligned}
\]
The corresponding equilibrium controls are
\[
\pi_t^*
=
\frac{h_t+\sigma_t Z_t+\sigma_t^0 Z_t^0}
{\gamma[\sigma_t^2+(\sigma_t^0)^2]},
\qquad
C_t^*
=
\delta^\psi\,\nu_t^{-\theta(\psi-1)}\,(\alpha e^{Y_t})^{-\frac{\psi}{\widetilde\theta}} X_t^*.
\]
A necessary stochastic maximum principle tailored to Epstein-Zin utility and a nonlinear transformation are the key ingredients in obtaining this characterization [2505.07231].

The stochastic maximum principle introduces adjoint processes \((p,q,q^0)\) and \(Q\), with Hamiltonian
\[
H(t,x,v,\pi,c,p,q,q^0,Q)
=
p[r_t x+\pi h_t-c]+(q\sigma_t+q^0\sigma_t^0)\pi-Qf(c\nu^{-\theta},v),
\]
and first-order conditions
\[
\partial_\pi H=0 \Longrightarrow p\,h_t+(q\sigma_t+q^0\sigma_t^0)=0,
\qquad
\partial_c H=0 \Longrightarrow p+Qf_1(c\nu^{-\theta},v)\nu^{-\theta}=0.
\]
The nonlinear transformation
\[
Y_t:=\ln p_t-\ln(\alpha U'(X_t^*))-\ln(-Q_t)
\]
then converts the adjoint system into the BSDE for equilibrium. According to the paper, this “log-ratio” transform is the key link between the SMP adjoints and the final BSDE characterization of the MFG Nash equilibrium [2505.07231].

In the deterministic case, where \(r,h,\sigma,\sigma^0\) depend only on \(t\), the BSDE reduces to a deterministic ODE and then to a Riccati equation, yielding an explicit closed-form solution for the equilibrium investment and consumption policies [2505.07231].

## 6. Solution concepts, convergence, and analytical themes

Across the supplied works, non-Markovian MFGs are solved through several distinct but related analytical mechanisms.

| Framework | Equilibrium device | Principal well-posedness or limit statement |
|---|---|---|
| Time-fractional subdiffusion [1812.05431] | Variational formulation via convex duality | Existence of a classical solution; uniqueness under monotonicity |
| Past absorptions [1902.02670] | Fixed point in feedback or relaxed feedback form | Existence of relaxed and strict feedback solutions; approximate Nash equilibria with \(\varepsilon_N\to0\) |
| Weak path-dependent MFG [2105.00484] | Generalised McKean-Vlasov BSDE | Unique solution in \(\mathcal{S}^\infty\times\mathcal{H}^2_{\mathrm{BMO}}\) under dissipativity and smallness or smoothness |
| Epstein-Zin portfolio game [2505.07231] | BSDE plus stochastic maximum principle | One-to-one correspondence between Nash equilibria and a class of BSDEs; uniqueness result |

In the time-fractional model, the coupled HJB-FP system is the Euler-Lagrange system of two dual convex problems. The primal functional \(\mathcal{A}(u)\) is defined on \(u\in C^2\) with \(u(T)=U_T\), the dual functional \(\mathcal{B}(m,w)\) is defined on \((m,w)\) satisfying the fractional-FP constraint, and Fenchel-Rockafellar duality yields
\[
\inf_u \mathcal{A}(u)=-\min_{(m,w)}\mathcal{B}(m,w).
\]
Taking first variations recovers precisely the HJB and FP equations with the correct fractional-time operators [1812.05431].

In the absorption model, existence proceeds by truncating drift and cost, applying the bounded-data existence result of Campi-Delarue via Kakutani’s theorem, passing to the limit using tightness in \(P_1(X)\) and a martingale-problem characterization, and finally using a mimicking theorem of Brunick-Shreve to convert the limit into feedback form [1902.02670].

In the weak non-Markovian path-space formulation, the equilibrium BSDE is used both for well-posedness and for limit theory. The paper proves non-asymptotic convergence estimates for the \(N\)-player game:
\[
|V^{i,N}-V^{\hat\xi}|^2
\le
C\Bigl(\frac1N+N R_N^2+\gamma_N\Bigr),
\]
and
\[
\int_0^T
W_2^2\Bigl(
\mathrm{Law}^{\mathbb{P}^{\hat\alpha^N}}(\hat\alpha_s^{i,N}),
\mathrm{Law}^{\mathbb{P}^{\hat\alpha}}(\hat\alpha_s)
\Bigr)\,ds
\le
C\Bigl(\frac1N+NR_N^2+\gamma_N\Bigr).
\]
Here \(R_N\to0\) measures the control-interaction correction and
\[
\gamma_N=
\sup_t \mathbb{E}[W_2^2(L^N(\mathbb{X}_t^N),\mu_t)]
+\mathbb{E}[W_2^2(L^N(\hat\alpha_t^N),\mathrm{Law}(\hat\alpha_t))].
\]
In the state-dependent-law case, the paper states that one can insert classical rates such as \(\gamma_N=O(N^{-1/2})\), leading overall to error of order \(O(N^{-1/2})\) in dimension \(m=1\) [2105.00484].

The same weak-form program also treats closed-loop equilibria and, in the Markovian case, the master equation. When a smooth master solution exists, the paper recovers
\[
Y_t=V(t,X_t,\mathrm{Law}(X_t))
\]
and establishes
\[
\mathbb{E}\bigl|v^{i,N}(0,\mathbb{X}_0^N)-V(0,X_0,\mathrm{Law}(X_0))\bigr|^2
=
O(N^{-1/2}),
\]
which yields convergence of the finite-\(N\) HJB system to the master equation [2105.00484].

A plausible implication of these results is that non-Markovianity shifts the analytic emphasis from PDE structure alone toward convex duality, weak formulations, stochastic maximum principles, and BSDE fixed points.

## 7. Conceptual distinctions, misconceptions, and research directions

One recurring distinction is between memory in dynamics and memory in interaction. In the time-fractional model, memory is generated at the microscopic level by a subdiffusion process \(X_t=Y_{E_t}\), and the macroscopic equations inherit Caputo derivatives [1812.05431]. In the absorption model, the state dynamics are standard diffusions but the mean-field interaction depends on survivor measures and the accumulated fraction absorbed, which is path-dependent through \((\mu_s)_{s\le t}\) and \(A_t\) [1902.02670]. In the weak formulation and Epstein-Zin portfolio game, non-Markovianity is instead encoded through full path dependence of coefficients, law dependence on path space, or recursive utility characterized by BSDEs [2105.00484; 2505.07231].

A second conceptual point concerns the role of monotonicity. In the time-fractional and absorption settings, uniqueness is obtained under monotonicity hypotheses of Lasry-Lions type [1812.05431; 1902.02670]. By contrast, the weak-form BSDE framework explicitly provides existence and uniqueness results that do not require Lasry-Lions monotonicity, replacing it with dissipativity of the drift and smallness or regularity assumptions on the terminal reward [2105.00484]. This does not eliminate monotonicity from the subject; rather, it identifies an alternative route to well-posedness in non-Markovian environments.

A third misconception is that abandoning Markov structure necessarily makes equilibrium characterization intractable. The supplied papers provide four counterexamples. Time-fractional systems preserve a variational Euler-Lagrange structure [1812.05431]. Absorption models admit strict and relaxed feedback fixed-point formulations and induce approximate Nash equilibria for large finite games [1902.02670]. Weak path-dependent games are characterized by generalized McKean-Vlasov BSDEs and support non-asymptotic convergence rates [2105.00484]. Recursive-utility portfolio games admit a one-to-one correspondence between Nash equilibria and BSDE solutions, and even explicit closed forms in the deterministic case [2505.07231].

The available research directions stated explicitly in the supplied material also point to broader generality. The Epstein-Zin study states that its combination of stochastic maximum principle, martingale-optimality principle, nonlinear adjoint-to-value transformation, and BSDE fixed-point arguments in BMO spaces can be exported to forward-utility MFGs, rank-dependent or habit-formation utilities, partial-information or jump-diffusion environments, and state-constraint or impulse-control MFGs [2505.07231]. This suggests that non-Markovian mean-field games are less a narrow subclass than a collection of methodologies for equilibrium analysis beyond the classical dynamic programming paradigm.

Source: https://www.emergentmind.com/topics/non-markovian-mean-field-games