---
title: Mean Field Games of Controls
url: https://www.emergentmind.com/topics/mean-field-games-of-controls
type: topic
---

# Mean Field Games of Controls

Mean field games of controls are mean field games in which a continuum of negligible agents interact through the joint distribution of their states and controls, rather than through the state distribution alone. In this setting, a representative player minimizes a cost functional that may depend on the law of state–control pairs, and equilibrium requires a fixed-point consistency between optimal feedback and the induced joint law. The literature treats this structure through forward–backward PDE systems, McKean–Vlasov FBSDEs, relaxed and measure-valued formulations, and numerical schemes adapted to nonlocal and often nonmonotone couplings [2508.21642] [2003.03968].

## 1. Historical emergence and defining feature

The general theory of mean field games studies deterministic or stochastic differential games in the limit as the number of agents tends to infinity. In the formulation emphasized in the numerical survey of control-interaction models, mean field games were introduced in the pioneering works of J.-M. Lasry and P.-L. Lions, and independently in the engineering literature by M. Y. Huang, P. E. Caines, and R. Malhamé [2003.03968].

The distinguishing feature of a mean field game of controls, often abbreviated MFGC, is that interaction occurs through the joint law of states and controls. One standard formulation uses a family of probability measures
$$
\mu_t \in \mathcal P(\Omega\times\mathbb R^n),
$$
with state marginal
$$
m_t := \mathrm{proj}_x{}_\#\mu_t \in \mathcal P(\Omega),
$$
so that the conditional law of the control given the state is encoded in $\mu_t$ [2508.21642]. Other formulations express the same idea through a local average control field $V(t,x)$, a price variable determined by aggregate control, or a conditional joint law under common noise [2003.03968] [1902.05461] [2205.13403].

This extension is not merely formal. It allows models in which agents react to average speed in a neighborhood, to aggregate production or price, or to the empirical distribution of strategies themselves. The data include crowd-motion models, debt refinance dynamics, pedestrian evacuation, exhaustible-resource production, price-interaction models, and flocking or synchronization examples [2003.03968] [2111.14209] [2405.01812] [2408.00733].

## 2. Canonical mathematical formulations

A representative-agent formulation on a bounded domain $\Omega\subset\mathbb R^n$ with control set $A=\mathbb R^n$ is given by
$$
dX_t = \alpha_t\,dt + \sqrt{2\nu}\,dB_t,\qquad X_0=x_0,
$$
together with the cost
$$
J^{x_0}(\alpha(\cdot);\mu)
=
\mathbb E\Big[
\int_0^T L(t,X_t,\alpha_t,\mu_t)\,dt
+\int_0^T f(t,X_t,m_t)\,dt
+g(X_T,m_T)
\Big].
$$
The Hamiltonian is the $p$-Legendre transform
$$
H(t,x,p,\mu):=\sup_{\alpha\in\mathbb R^n}[p\cdot\alpha-L(t,x,\alpha,\mu)],
$$
and equilibrium is described by the forward–backward system
$$
\begin{cases}
-\,u_t-\nu\Delta_x u + H(t,x,D_xu,\mu_t)=f(t,x,m(t)),\\[0.5ex]
m_t-\nu\Delta_x m-\nabla_x\!\cdot\!\bigl(m\,D_pH(t,x,D_xu,\mu_t)\bigr)=0,\\[0.5ex]
\mu_t=\bigl(\mathrm{id},-D_pH(t,\cdot,D_xu(t,\cdot),\mu_t)\bigr)\#\,m_t,\\[0.5ex]
u(T,x)=g(x,m(T)),\quad m(0,x)=m_0(x).
\end{cases}
$$
This is the basic PDE realization of MFGC on bounded domains [2508.21642].

A closely related Dirichlet formulation on a closed domain writes the dynamics as
$$
dX_t=b(t,X_t,\alpha_t;\mu_t)\,dt+2\sqrt{\sigma}\,dW_t,
$$
stopped upon exit, with Hamiltonian
$$
H(t,x,p;\mu)=\sup_{\alpha\in A}\{-\,p\cdot b(t,x,\alpha;\mu)-\mathcal L(t,x,\alpha;\mu)\},
$$
and fixed-point condition
$$
\mu_t=\bigl(\mathrm{Id},\alpha^*(t,\cdot,Du;\mu_t)\bigr)_\#\,m_t.
$$
The induced PDE system couples HJB and Fokker–Planck equations under absorbing boundary conditions [2111.14209].

The probabilistic side is equally prominent. In the Pontryagin formulation of large-population games with interaction through the empirical distribution of states and controls, equilibria satisfy McKean–Vlasov FBSDEs. In common-noise settings the master equation is posed on the joint law of state and control, and the population is represented by a random probability measure $\nu_t\in\mathcal P_2(\mathbb R^{2d})$ with state marginal $\mu_t=\pi_\#^1\nu_t$ [2004.08351] [2205.13403]. In a Hilbert-space formulation, the equilibrium control is characterized as the zero of a monotone variational inequality built from a coupled mean-field FBSDE [2602.14621].

This suggests that MFGC is less a single model than a family of equivalent equilibrium formalisms—PDE, FBSDE, martingale problem, and variational inequality—linked by the same fixed-point constraint on the joint law of state and control.

## 3. Boundary-value problems and constrained-state variants

Boundary conditions are a central structural feature in recent MFGC analysis. On bounded domains, two cases are treated in parallel: Dirichlet boundary conditions, interpreted as absorption or exhaustion, and Neumann boundary conditions, interpreted as reflection [2508.21642].

For the system above, the boundary conditions on $\Sigma=[0,T]\times\partial\Omega$ are
- Dirichlet:
  $$
  u=0,\qquad m=0;
  $$
- Neumann:
  $$
  \partial_n u=0,\qquad \nu\partial_n m + m\,D_pH\cdot n=0.
  $$

In the Dirichlet case, the condition $m|_{\partial\Omega}=0$ models absorption at exit, and mass is not conserved. This lack of mass conservation is the principal technical novelty in the Dirichlet analysis of Bongini and Salvarani, which develops a priori estimates specifically adapted to the fact that standard $L^1$ and maximum-principle arguments cannot be used directly [2111.14209]. By contrast, the Neumann condition yields mass preservation and a no-flux boundary relation, and in the numerical reflected-diffusion formulation it arises from Itô’s formula applied to
$$
dX_t=\alpha_t\,dt+\sqrt{2\nu}\,dB_t-2\nu\,n(X_t)\,dL_t
$$
with reflection on $\partial\Omega$ [2003.03968].

Reflection can also be imposed against a stochastic boundary process. In the reflected-state model with an exogenous continuous boundary $A_t$, the state solves
$$
dX_t=b(t,X_t,\mu_t,a_t)\,dt+\sigma(t,X_t,\mu_t,a_t)\,dW_t+dR_t,
$$
where $R$ is continuous, nondecreasing, $X_t\ge A_t$, and
$$
\int_0^T \mathbf 1_{\{X_t>A_t\}}\,dR_t=0.
$$
The corresponding dynamic Skorokhod problem yields a Lipschitz Skorokhod map $\Gamma$ and permits a weak relaxed equilibrium theory on an enlarged canonical space [2503.03253].

State constraints also appear in application-specific forms. In the Cournot exhaustible-resource model, the state is an inventory level on $(0,L)$ with reflection at $x=L$ and absorption at $x=0$; the HJB uses $u(t,0)=0$ and $\partial_xu(t,L)=0$, while the Fokker–Planck equation uses $m(t,0)=0$ and a reflecting boundary relation at $L$ [2405.01812]. Across these examples, boundary conditions are not secondary regularity choices; they encode exit, depletion, preservation of mass, or exogenous state reflection.

## 4. Existence, uniqueness, monotonicity, and nonuniqueness

Two main well-posedness mechanisms are explicit in the recent PDE theory on bounded domains. The first is a small-coupling theory without monotonicity. Under coercivity and controlled growth of $H$ and $D_pH$, Lipschitz and Hölder dependence on the measure argument, uniformly bounded terminal cost $g$ in $C^{3+\beta}$, and $f\equiv 0$, a quantitative smallness condition on the constants in the growth and convexity bounds yields a unique classical solution $(u,m,\mu)$ for both Dirichlet and Neumann boundary conditions [2508.21642].

The second is a monotone-coupling theory in the Lasry–Lions sense. If $L(t,x,\alpha,\mu)$ is strictly convex in $\alpha$, $C^1$ in $(x,\alpha)$, and satisfies
$$
\int \bigl(L(t,x,\alpha,\mu_1)-L(t,x,\alpha,\mu_2)\bigr)\,d(\mu_1-\mu_2)\ge 0,
$$
while the state couplings $f$ and $g$ are monotone in $m$, then for arbitrary $T>0$ there exists a unique strong solution $(u,m,\mu)$ with either Dirichlet or Neumann boundary conditions [2508.21642]. A related theory on $\mathbb R^d$ and on the torus assumes power-type growth of the Hamiltonian and monotonicity in the law of states and controls, and proves existence and uniqueness through a priori estimates and Leray–Schauder [2006.12949].

The Dirichlet theory of Bongini and Salvarani employs a double fixed-point procedure: first solve the fixed point
$$
\mu_t=\bigl(\mathrm{Id},\alpha^*(Du(t,\cdot),\mu_t)\bigr)_\#m_t,
$$
then solve linear HJB and Fokker–Planck equations, and finally close the map by Leray–Schauder. Much of the argument is devoted to positivity and mass estimates for $m$ despite absorbing boundaries [2111.14209]. In the 2025 boundary-value treatment, the proof strategy combines Leray–Schauder, a priori $L^\infty$ and $L^q$ estimates for $u$ and $D_xu$, a Bernstein-type boundary-adapted argument for $\|D_xu\|_\infty$, and Moser iteration or classical parabolic theory for $m$ [2508.21642].

Monotonicity is also studied at the master-equation level. Mou and Zhang analyze Lasry–Lions monotonicity, displacement monotonicity, anti-monotonicity, and displacement $\lambda$-monotonicity or semi-monotonicity for MFGC with common noise, and prove that these properties propagate along classical master flows. They describe this as the first step toward a global well-posedness theory for master equations of mean field games of controls [2205.13403].

A frequent misunderstanding is that interaction through controls admits either complete uniqueness or complete instability. The literature is more differentiated. The finite-difference study stresses that in models where agents try to adjust speed to an average speed, the monotonicity assumptions frequently made in MFG theory do not hold and uniqueness cannot be expected in general; its symmetric numerical example exhibits at least three distinct solutions [2003.03968]. At the same time, uniqueness is recovered under monotone couplings [2508.21642], under quantitative smallness assumptions [2508.21642], and even without monotonicity in a linear–quadratic setting with common noise, where the Nash certainty equivalence system and the associated master equation are solvable over an arbitrary time horizon [2206.01732].

## 5. Finite-player limits, relaxed equilibria, and singular controls

A substantial part of the MFGC literature concerns the relation between finite $N$-player games and the mean-field limit. In the open-loop framework with interaction through the empirical distribution of states and controls, Laurière and Tangpi derive convergence of Nash equilibria to the unique mean field equilibrium by developing propagation of chaos for forward and backward weakly interacting particles. Their results include moment and concentration bounds, and in linear–quadratic examples the rate is $O(1/N)$ [2004.08351].

A complementary route uses controlled Fokker–Planck equations and measure-valued equilibria. In that framework, $\epsilon_N$-Nash equilibria in $N$-player games have limits as $N\to\infty$, and each limit is a measure-valued solution of the mean field game of controls; conversely, any measure-valued solution can be obtained as the limit of a sequence of $\epsilon_N$-Nash equilibria. The same work also proves existence of measure-valued solutions in the case without common noise [2006.12993].

For reflected states and joint-law dependence, Bo, Wang, and Yu establish a relaxed mean field equilibrium on an enlarged canonical space built around the dynamic Skorokhod mapping. They further prove two directional limit results: any weak limit of empirical measures induced by $\boldsymbol\epsilon$-Nash equilibria in $N$-player games is supported on the set of relaxed mean field equilibria, and any Markovian mean field equilibrium in the weak sense can be approximated by a sequence of constructed $\boldsymbol\epsilon$-Nash equilibria as $N\to\infty$ [2503.03253].

Singular control enters the MFGC literature through relaxed formulations and compactness in the Skorokhod weak $M_1$ topology. In the extended mean-field game with interaction through both states and controls, relaxed regular controls are combined with singular finite-variation controls, and existence is proved by first smoothing singular controls into continuous controls, establishing compactness and closed-graph properties of the best-response correspondence, and then passing to the singular limit in weak $M_1$ topology [1909.04154]. An approximation result for $N$-player stochastic games with singular controls shows that the optimal control of an MFG with bounded velocity $\theta$ is an $\epsilon_N$-Nash equilibrium of the bounded-velocity $N$-player game with
$$
\epsilon_N = O\!\bigl(N^{-1/2}\bigr),
$$
and an $(\epsilon_N+\epsilon_\theta)$-Nash equilibrium of the finite-variation game, with $\epsilon_\theta\to0$ as $\theta\to\infty$ [2202.06835].

## 6. Potential structures, linear–quadratic classes, and nonlocal extensions

One major subclass of MFGC is potential. On the torus, a market-price model with endogenous price
$$
P(t)=\Psi\!\Bigl(t,\int_{\mathbb T^d}\phi(x,t)\,v(x,t)\,m(x,t)\,dx\Bigr)
$$
leads to the coupled system
$$
\begin{cases}
-\partial_tu-\sigma\Delta u + H(x,t,\nabla u+\phi^\top P)=f(x,t,m),\\
\partial_t m-\sigma\Delta m+\mathrm{div}(v\,m)=0,\\
v=-H_p(x,t,\nabla u+\phi^\top P).
\end{cases}
$$
Here an incomplete potential and, under an additional monotonicity-of-$f$ assumption, a full potential functional $\mathcal B(m,v)$ are available, and existence of a unique classical solution follows from Leray–Schauder together with Schauder estimates [1902.05461].

A broader non-Markovian perspective appears in the mean field control approach of Höfer and Soner. They show that minimizers of a mean field optimal control problem with common noise and jumps are Nash equilibria of an associated mean field game of controls, and that these associated games are necessarily potential. Their examples include a mean field game of controls with interactions through a price variable, and mean field Cucker–Smale flocking and Kuramoto models [2408.00733].

The linear–quadratic class is especially explicit. In the common-noise model with state
$$
dX_t=\bigl[A\,X_t+B\,\alpha_t+f(\nu_t)+b(\mu_t)\bigr]dt+\sigma\,dW_t+\sigma_0\,dW^0_t,
$$
the stochastic maximum principle yields a feedback control
$$
\alpha_t^*=-R^{-1}B\bigl(P_tX_t+\Lambda_t\bigr)-h(\mu_t),
$$
the mean field equilibrium is determined by a Nash certainty equivalence system, and the master equation reduces to a finite-dimensional second-order parabolic equation. The paper proves global existence and uniqueness of the master equation on arbitrary time horizons without monotonicity conditions, and quantitative convergence of the $N$-player game to the mean field game with $O(N^{-1})$ estimates [2206.01732].

Nonlocal diffusion has also entered the theory. In the fractional MFGC on $\mathbb T^d$ with $s\in(\tfrac12,1)$, the Brownian Laplacian is replaced by the generator of a $2s$-stable pure-jump process,
$$
(-\Delta)^s\phi(x)=\sum_{k\in\mathbb Z^d}(2\pi|k|)^{2s}\hat\phi_k\,e^{2\pi ik\cdot x},
$$
and existence is proved under Lasry–Lions monotonicity by combining moment estimates on $\mu$, abstract fractional-parabolic estimates, and time-regularity estimates for the joint law derived from the associated Lévy process [2509.04647].

## 7. Computation, learning, and model behavior

The computational literature reflects the dual difficulty of MFGC systems: they are forward–backward and nonlocal, and they may be nonmonotone. A finite-difference approximation on bounded domains discretizes the HJB backward in time and the Fokker–Planck equation forward in time, uses a monotone Hamiltonian built from an upwind Godunov flux, enforces discrete Neumann conditions through ghost nodes, and preserves nonnegativity and total mass at the discrete level in the reflecting case [2003.03968].

The associated nonlinear solver combines continuation, Newton iterations, and an inner bigradient-like loop. In the reported experiments, Newton has local quadratic convergence, in practice $3$–$6$ outer steps suffice, BiCGStab takes on average $5$–$40$ inner iterations, and the method captures both “gathering–kissing–splitting” in a square and smooth queue formation in a corridor [2003.03968]. The same study emphasizes multiplicity in symmetric nonmonotone regimes, which is consistent with the absence of a general uniqueness theory outside monotone or specially structured classes [2003.03968].

For Cournot mean field games of controls, a learning-based numerical method is built around Smoothed Policy Iteration. Given a smooth policy guess, one alternates Fokker–Planck policy evaluation, price update, HJB policy evaluation, best response, and smoothing
$$
\bar q^{(n+1)}=(1-\zeta_n)\bar q^{(n)}+\zeta_n q^{(n+1)},\qquad \zeta_n=\beta/(n+\beta).
$$
The paper proves uniqueness of the equilibrium under general assumptions on the price function and proves convergence of the learning algorithm; numerically it reproduces depletion dynamics and the Hubbert peak in oil-production models [2405.01812].

A different numerical perspective treats MFGC equilibria as zeros of a monotone variational inequality in a Hilbert space. The extragradient iteration
$$
\alpha_{n+\frac12}=\alpha_n-\gamma\,v(\alpha_n),\qquad
\alpha_{n+1}=\alpha_n-\gamma\,v(\alpha_{n+\frac12})
$$
is shown to converge with $O(1/n)$ averaged rate under strong monotonicity and Lipschitz assumptions, and exponentially fast for the last iterate under stronger assumptions. The paper explicitly relates this construction to fictitious play and extends it to general mean field type FBSDEs that do not necessarily come from optimal control [2602.14621].

Taken together, these methods indicate that computation in mean field games of controls is not tied to a single numerical paradigm. Finite differences, policy iteration, continuation–Newton methods, and monotone-operator algorithms all appear, and the choice depends on whether the target problem is boundary-driven, potential, monotone, nonmonotone, or formulated primarily through PDEs or FBSDEs.

Source: https://www.emergentmind.com/topics/mean-field-games-of-controls