NPMAI ECOSYSTEM · Research Mathematics Department · 2026
Mathematics · Stochastic Control · Mean Field Games · 2026

A Generalized Linear-Quadratic Mean Field Game with Common Noise, Mean-Reverting Dynamics, and Exogenous Tracking

Hierarchical Affine Decoupling and an Explicit Decentralized Nash Equilibrium via Coupled Riccati Systems

NPMAI ECOSYSTEM · About the Organisation

An open research & development organisation, publishing original work across departments.

NPMAI ECOSYSTEM is a multi-domain research and engineering organisation based in Kota, Rajasthan. It runs production AI tooling used by hundreds of thousands of developers, alongside an original research programme spanning computer science, social science, chemistry, and mathematics — the department responsible for the paper that follows.

3.5M+
npmai downloads
45+
LLM endpoints
1,371
agent tools
4
research departments
$0
monthly infrastructure cost

Research Tech (CS, AI)

Agentic systems, RAG architectures, load balancing, and applied ML infrastructure.

Social Science & Legal

Political science frameworks and constitutional law proposals, grounded in real governance data.

Chemistry Research

Reactor design, catalysis, and synthetic fuel pathways — including hydrogen and materials work.

Mathematics Research

Stochastic control, game theory, and dynamical systems. Home department of this paper.

Sonu Kumar, Founder, NPMAI ECOSYSTEM

Sonu Kumar

Founder, NPMAI ECOSYSTEM · HOD, Research Tech (CS, AI) & Social Science (Political Science, Legal) · "Bihar Viral Boy"

Sonu Kumar founded NPMAI ECOSYSTEM and built its foundational infrastructure — the npmai package, the dual load balancer, and the LARA retrieval architecture — entirely on free-tier cloud infrastructure. He leads the organisation's Research Tech and Social Science departments and set up the multi-department research programme under which this Mathematics Department paper is published. He is not an author of the present paper.

Founder TEDx Speaker Age 15 Kota, Rajasthan Allen Career Institute — full scholarship
Dr. Brajesh Maheshwari
Dr. Brajesh Maheshwari
HOD, Research Physics
Aadarsh Singh
Aadarsh Singh
HOD, Research Mathematics — author of this paper
Sampatirao Trilok Varma
Sampatirao Trilok Varma
HOD, Research Chemistry
Read the research paper
Abstract

Classical continuous-time Linear-Quadratic Mean Field Games (LQ-MFGs) typically restrict attention to simple linear drifts, quadratic control costs, and interactions that depend only on the population average. This paper develops a generalized LQ-MFG that simultaneously incorporates four mechanisms absent from the classical formulation: asymmetric Ornstein–Uhlenbeck mean-reverting dynamics, a stochastic common noise shared by the entire population, a deterministic exogenous tracking trajectory M(t), and a control–state interaction penalty coupling individual control effort to the aggregate state.

Using the Stochastic Maximum Principle, we derive the coupled Forward-Backward Stochastic Differential Equation (FBSDE) governing the representative agent, and — through a hierarchical affine decoupling ansatz applied first at the macroscopic (population) level and then at the microscopic (individual) level — reduce the coupled stochastic system to two nonlinear Riccati differential equations and two linear ordinary differential equations. We solve these explicitly in closed form and, invoking a Mean Field Consistency condition, obtain a fully explicit decentralized Nash equilibrium control law. The equilibrium decomposes naturally into a local consensus-correction term, a system-wide stabilization term generated by the common noise and the control–state coupling, and an anticipatory tracking term generated by the exogenous trajectory. Several existing LQ-MFG models are recovered as special cases.

Keywords — Mean Field Games, Linear-Quadratic Control, Common Noise, Ornstein–Uhlenbeck Dynamics, Stochastic Maximum Principle, Forward-Backward SDE, Riccati Equation, Nash Certainty Equivalence, Exogenous Tracking, Decentralized Equilibrium

01 System Model and Mathematical Framework

1.1 Introduction and Motivation

Mean Field Games (MFGs), introduced independently by Jean-Michel Lasry and Pierre-Louis Lions, provide a mathematical framework for analyzing strategic interactions among an extremely large population of rational agents. Rather than solving an intractable finite-player dynamic game, the influence of each individual agent is assumed to vanish as the population size approaches infinity, allowing every agent to optimize against an aggregate statistical description of the population known as the mean field. This limiting formulation has found applications in economics, finance, engineering, networked control systems, opinion dynamics, smart grids, and large-scale autonomous systems.

Among the numerous classes of Mean Field Games, Linear-Quadratic Mean Field Games (LQ-MFGs) occupy a particularly important position because they combine analytical tractability with the ability to model complex interactions through linear dynamics and quadratic objective functionals. Their mathematical structure frequently permits explicit equilibrium characterizations through Forward-Backward Stochastic Differential Equations (FBSDEs) and Riccati differential equations, making them valuable both theoretically and computationally.

Classical LQ-MFG formulations generally assume relatively simple state dynamics in which each agent evolves under a linear drift, quadratic control costs, and interactions that depend solely upon the population average. Although these assumptions simplify the equilibrium analysis, they neglect several important phenomena frequently encountered in practical systems.

First, many physical, economic, and biological systems exhibit mean-reverting behavior, whereby an individual's state naturally tends toward an equilibrium level unless continuously acted upon. Such dynamics are well represented by Ornstein–Uhlenbeck processes and arise naturally in financial interest-rate models, inventory regulation, energy markets, and population dynamics.

Second, agents often attempt to track not only the collective population behavior but also an externally imposed reference trajectory. Such trajectories may represent macroeconomic indicators, regulatory benchmarks, technological targets, environmental indices, or desired production schedules. Incorporating these time-varying reference signals significantly enriches the strategic behavior of the game while preserving its linear-quadratic structure.

Third, real-world systems are frequently exposed to common sources of uncertainty. Unlike independent idiosyncratic disturbances, common noise simultaneously affects every participant and therefore remains present even in the infinite-population limit. Consequently, the mean field itself becomes stochastic rather than deterministic, requiring the equilibrium analysis to be performed conditionally with respect to the filtration generated by the common noise.

Motivated by these considerations, this paper develops a generalized continuous-time Linear-Quadratic Mean Field Game that simultaneously incorporates asymmetric mean-reverting dynamics, stochastic common noise, an exogenous tracking trajectory, and an additional interaction penalty coupling the individual control effort with the aggregate state. The resulting optimization problem remains analytically tractable while substantially extending the class of models that admit explicit closed-form equilibria.

The principal contributions of this work may be summarized as follows:

  1. We formulate a generalized asymmetric Linear-Quadratic Mean Field Game incorporating Ornstein–Uhlenbeck-type mean-reverting dynamics with cross-population coupling.
  2. We introduce a deterministic, time-varying exogenous tracking trajectory into the running cost without destroying the analytical solvability of the model.
  3. We incorporate a control–state interaction penalty that modifies the aggregate dynamics and produces an additional feedback component in the equilibrium strategy.
  4. Using the Stochastic Maximum Principle, we derive the associated coupled Forward-Backward Stochastic Differential Equation governing the representative agent.
  5. By employing a hierarchical affine decoupling ansatz, we reduce the coupled stochastic system to a collection of nonlinear Riccati differential equations together with non-homogeneous linear ordinary differential equations.
  6. Finally, we obtain an explicit decentralized Nash equilibrium control law expressed entirely in closed form, providing a complete analytical characterization of the equilibrium.

The remainder of the paper is organized as follows. Section 1 establishes the mathematical model and the stochastic dynamics governing the population. Section 2 derives the representative agent's optimization problem and the associated FBSDE. Section 3 develops the analytical solution through a hierarchical decoupling procedure and states the explicit decentralized Nash equilibrium. Section 4 records structural corollaries and special cases. Section 5 concludes.

1.2 Mathematical Framework and Notation

Throughout this paper, we consider a finite time horizon $T>0$ defined on a complete filtered probability space

$$(\Omega,\mathcal{F},\{\mathcal{F}_t\}_{0\le t\le T},\mathbb{P}),$$

satisfying the usual completeness and right-continuity assumptions. Unless otherwise stated, all stochastic processes are assumed adapted to $\{\mathcal{F}_t\}$. The game consists of an infinite continuum of strategically interacting agents indexed by $i\in[0,1]$. Each agent controls a one-dimensional stochastic state process $X_i(t)\in\mathbb{R}$, whose evolution depends upon its own control action, the aggregate population behavior, and both idiosyncratic and common sources of uncertainty.

SymbolDescription
$X_i(t)$State of agent $i$.
$\bar X(t)$Conditional population mean field.
$u_i(t)$Admissible control process of agent $i$.
$W_i(t)$Idiosyncratic Brownian motion.
$W_0(t)$Common Brownian motion.
$a$Mean-reversion coefficient.
$b$Mean-field interaction coefficient.
$\sigma$Idiosyncratic volatility intensity.
$\sigma_0$Common volatility intensity.
$M(t)$Deterministic tracking trajectory.
$X_D$Terminal target state.
$q,\eta,g,\gamma$Cost-function parameters.

The admissible control space is defined by

$$\mathcal U=\left\{u:\Omega\times[0,T]\rightarrow\mathbb R \;\middle|\; u\text{ progressively measurable, } \mathbb E\!\left[\int_0^T |u(t)|^2dt\right]<\infty\right\},$$

ensuring that every admissible strategy possesses finite expected energy and yields a well-defined stochastic state trajectory.

1.3 Model Formulation

We consider an infinite population of atomless strategic agents whose states evolve over the finite planning horizon $t\in[0,T]$. Each agent seeks to regulate its own state while simultaneously responding to the aggregate population behavior and an externally prescribed reference trajectory. Unlike classical LQ-MFGs, the present formulation incorporates four structural mechanisms simultaneously: intrinsic mean-reverting dynamics, asymmetric interaction with the population mean, stochastic common environmental shocks, and exogenous trajectory tracking.

The individual state dynamics are governed by the stochastic differential equation

$$dX_i(t)=\big(-aX_i(t)+b\bar X(t)+u_i(t)\big)\,dt+\sigma\,dW_i(t)+\sigma_0\,dW_0(t),$$

where $a>0$ denotes the intrinsic rate of mean reversion and $b\in\mathbb R$ measures the influence of the aggregate population on the individual dynamics. Positive values of $b$ encourage conformity with the population average, whereas negative values correspond to repulsive interactions. The process $W_i(t)$ represents idiosyncratic uncertainty unique to each agent, while $W_0(t)$ models common environmental uncertainty simultaneously experienced by the entire population, with diffusion intensities $\sigma$ and $\sigma_0$ respectively.

The aggregate population state is defined by the conditional expectation

$$\bar X(t)=\mathbb E\big[X_i(t)\mid \mathcal F_t^0\big],\qquad \mathcal F_t^0=\sigma\big(W_0(s),\,0\le s\le t\big),$$

the filtration generated exclusively by the common Brownian motion. This conditional formulation is fundamental to Mean Field Games with common noise: since the independent idiosyncratic Brownian motions satisfy an exact law of large numbers over the continuum of agents, their aggregate contribution vanishes in the infinite-population limit. Only the common noise survives, rendering the mean field itself a stochastic process adapted to the common filtration.

Each player chooses an admissible control $u_i\in\mathcal U$ so as to minimize the quadratic performance functional

$$J_i(u_i)=\mathbb E\left[\int_0^T\left(\tfrac12u_i^2+\tfrac q2(X_i-\bar X)^2+\gamma u_i\bar X+\tfrac\eta2(X_i-M(t))^2\right)dt+\tfrac g2\big(X_i(T)-X_D\big)^2\right].$$

The objective functional consists of four components: a quadratic control cost discouraging excessive control effort; a consensus penalty promoting alignment between an individual's state and the population average, modeling herding behavior; a mixed control–state interaction term capturing situations where the effectiveness or expense of control depends on the aggregate state; and a tracking term driving each agent toward the prescribed trajectory $M(t)$, with a terminal cost enforcing convergence toward the desired terminal target $X_D$.

Agent i X_i(t) Mean Field X̄(t) Tracking M(t), X_D Control u_i(t) feedback into dynamics (b, γ, σ₀)
Fig. 1 — Each agent's control responds to its own state, the stochastic mean field, and the exogenous tracking trajectory; the resulting controls jointly regenerate the mean field.

The interaction of these four mechanisms produces a generalized Linear-Quadratic Mean Field Game whose equilibrium structure is characterized analytically in the following sections through the Stochastic Maximum Principle and the associated coupled Forward-Backward Stochastic Differential Equations.

02 Decentralized Optimality and the Coupled Forward-Backward Stochastic System

Having established the stochastic dynamics and objective functional governing the representative agent, we now derive the optimal decentralized strategy. Since the game consists of an infinite continuum of strategically interacting players, the influence of any single agent on the aggregate population state becomes negligible in the mean-field limit. Consequently, each player treats the population average $\bar X(t)$ as an exogenously given stochastic process while solving an individual stochastic optimal control problem.

This principle, commonly referred to as the Nash Certainty Equivalence (NCE) Principle, forms the foundation of continuous-time Mean Field Game theory. Each agent computes an optimal response to the anticipated evolution of the mean field, and equilibrium is achieved only when the resulting population behavior reproduces the same mean field that was originally assumed. Throughout this section, we regard $\bar X(t)$ as a known $\mathcal F_t^0$-adapted stochastic process; the consistency requirement is imposed in Section 3.

2.1 Representative Agent Optimization Problem

Fix an arbitrary agent $i$. Its state dynamics satisfy

$$dX(t)=\big(-aX(t)+b\bar X(t)+u(t)\big)dt+\sigma\,dW(t)+\sigma_0\,dW_0(t),\qquad X(0)=X_0,$$

with $u(t)$ progressively measurable and of finite second moment. The objective is to minimize

$$J(u)=\mathbb E\left[\int_0^T L(t,X,u,\bar X)\,dt+\Phi(X(T))\right],$$

where the running cost is

$$L(t,X,u,\bar X)=\tfrac12u^2+\tfrac q2(X-\bar X)^2+\gamma u\bar X+\tfrac\eta2(X-M(t))^2,$$

and the terminal penalty is $\Phi(X(T))=\tfrac g2(X(T)-X_D)^2$. The quadratic structure guarantees convexity of the optimization problem with respect to the control variable, ensuring uniqueness of the optimal decentralized strategy.

2.2 Application of the Stochastic Maximum Principle

To characterize optimal controls, we employ the Stochastic Maximum Principle, converting the stochastic optimization problem into a coupled FBSDE. Introduce the adjoint process $p(t)$ — the marginal value of the state variable — together with the martingale correction processes $z(t)$ and $z_0(t)$, associated respectively with the idiosyncratic and common Brownian motions. The stochastic Hamiltonian is

$$H(t,x,u,p)=p\big(-ax+b\bar X+u\big)+\tfrac12u^2+\tfrac q2(x-\bar X)^2+\gamma u\bar X+\tfrac\eta2(x-M(t))^2.$$

The first term measures the sensitivity of the future value function to the current dynamics, while the remaining terms represent the instantaneous running cost.

2.3 First-Order Optimality Condition

Since $H$ is strictly convex in $u$, differentiating gives

$$\frac{\partial H}{\partial u}=p+u+\gamma\bar X = 0 \quad\Longrightarrow\quad \boxed{u^*(t)=-p(t)-\gamma\bar X(t).}$$

The optimal policy separates into the classical adjoint feedback $-p(t)$ and an additional correction $-\gamma\bar X(t)$ generated by the control–state interaction penalty. When $\gamma=0$, the model reduces to the standard LQ-MFG feedback law.

2.4 The Adjoint Backward Equation

The adjoint process evolves according to

$$dp(t)=-\frac{\partial H}{\partial x}\,dt+z(t)\,dW(t)+z_0(t)\,dW_0(t).$$

Since $\partial H/\partial x = -ap+q(x-\bar X)+\eta(x-M)$, the adjoint system is

$$dp(t)=\Big(ap(t)-q(X(t)-\bar X(t))-\eta(X(t)-M(t))\Big)dt+z(t)dW(t)+z_0(t)dW_0(t),$$ $$p(T)=g\big(X(T)-X_D\big).$$

2.5 The Coupled Forward-Backward System

Substituting $u^*(t)=-p(t)-\gamma\bar X(t)$ into the state dynamics and collecting terms yields the forward equation $dX=\big(-aX+(b-\gamma)\bar X-p\big)dt+\sigma\,dW+\sigma_0\,dW_0$. Together with the adjoint equation, the representative agent satisfies the coupled FBSDE:

$$\begin{cases} dX=\big(-aX+(b-\gamma)\bar X-p\big)dt+\sigma\,dW+\sigma_0\,dW_0,\\[4pt] dp=\big(ap-q(X-\bar X)-\eta(X-M)\big)dt+z\,dW+z_0\,dW_0,\\[4pt] X(0)=X_0,\qquad p(T)=g\big(X(T)-X_D\big). \end{cases}$$

This system constitutes the mathematical core of the equilibrium problem: the forward equation governs the controlled state, while the backward equation propagates the shadow value of the state from the terminal time back to the initial time. Their mutual dependence through the control creates a fully coupled stochastic boundary-value problem.

Existence of an Optimal Control. Under the assumptions of Section 1, the drift and diffusion coefficients are globally Lipschitz, the running cost is uniformly convex in the control, and the terminal cost is convex. Standard results from stochastic optimal control guarantee the existence of a unique adapted solution to the representative optimization problem, provided the associated FBSDE admits a unique adapted solution — constructed explicitly in Section 3 through an affine decoupling ansatz.

03 Analytical Construction of the Nash Equilibrium

The coupled FBSDE derived above constitutes a nonlinear stochastic boundary-value problem. Although such systems generally do not admit explicit analytical solutions, the linear-quadratic structure of the present model permits a complete reduction to deterministic differential equations through an appropriate affine decoupling procedure. Rather than solving the full coupled system directly, we proceed hierarchically: first the macro-level problem, governing the evolution of the population average, then the micro-level problem, describing the optimal response of a representative agent conditional on the macro solution.

Coupled Stochastic FBSDE (X, p) Macro layer — §3.1 Riccati φ(t) + linear ψ(t) Micro layer — §3.2 Riccati Φ(t), Θ(t), Ξ(t) Consistency: Φ+Θ=φ, Ξ=ψ — §3.3 Explicit Decentralized Nash Equilibrium (Theorem 3.1)
Fig. 2 — Hierarchical affine decoupling: the coupled stochastic FBSDE reduces to two deterministic Riccati systems, joined by the Mean Field consistency condition.

3.1 Aggregate Mean Field Dynamics and Macro-Level Decoupling

Averaging the optimal individual dynamics eliminates the idiosyncratic Brownian motions through the exact law of large numbers, leaving only the common source of randomness. The optimal control from Section 2 is $u_i^*(t)=-p_i(t)-\gamma\bar X(t)$. Taking conditional expectations with respect to $\mathcal F_t^0$ gives the aggregate optimal control $\bar u(t)=-\bar p(t)-\gamma\bar X(t)$, where $\bar p(t)=\mathbb E[p_i(t)\mid\mathcal F_t^0]$. Substituting into the averaged state equation and defining the effective macro drift coefficient $\alpha=-a+b-\gamma$, the aggregate dynamics become

$$d\bar X(t)=\big(\alpha\bar X(t)-\bar p(t)\big)dt+\sigma_0\,dW_0(t).$$

Unlike classical LQ-MFGs without common noise, the mean field is itself stochastic because the common Brownian motion survives averaging. Since $\mathbb E[X-\bar X\mid\mathcal F_t^0]=0$, the consensus penalty disappears after conditional expectation, giving the aggregate backward equation

$$d\bar p(t)=\big(a\bar p(t)-\eta(\bar X(t)-M(t))\big)dt+\bar z_0(t)\,dW_0(t),\qquad \bar p(T)=g(\bar X(T)-X_D).$$

Affine decoupling ansatz. We postulate $\bar p(t)=\phi(t)\bar X(t)+\psi(t)$, where $\phi(t)$ is a deterministic feedback gain and $\psi(t)$ a deterministic tracking bias. Applying Itô's formula and matching coefficients against the aggregate backward equation (using uniqueness of the semimartingale decomposition) yields:

$$\dot\phi+(\alpha-a)\phi-\phi^2+\eta=0,\qquad \phi(T)=g,$$ $$\dot\psi-(a+\phi)\psi=-\eta M(t),\qquad \psi(T)=-gX_D,$$ $$\bar z_0(t)=\phi(t)\sigma_0.$$

The original stochastic macro problem is thus reduced to a nonlinear Riccati equation for $\phi(t)$ and a non-homogeneous linear equation for $\psi(t)$.

3.2 Individual Tracking Layer and the Micro-Level Riccati System

Because every coefficient in the individual FBSDE is linear in both $X(t)$ and $\bar X(t)$, we seek the affine ansatz $p(t)=\Phi(t)X(t)+\Theta(t)\bar X(t)+\Xi(t)$, where $\Phi(t)$ measures sensitivity to the agent's own state, $\Theta(t)$ to the aggregate state, and $\Xi(t)$ is the deterministic tracking component. Applying Itô's formula, substituting the forward equations for $X$ and $\bar X$, and matching against the backward equation term-by-term gives:

$$\dot\Phi-2a\Phi-\Phi^2+(q+\eta)=0,\qquad \Phi(T)=g,$$ $$\dot\Theta-(a+\Phi-\alpha+\phi)\Theta=q-\Phi(b-\gamma),\qquad \Theta(T)=0,$$ $$\dot\Xi-(a+\Phi)\Xi=-\Theta\psi+\eta M,\qquad \Xi(T)=-gX_D,$$ $$z(t)=\Phi(t)\sigma,\qquad z_0(t)=\big(\Phi(t)+\Theta(t)\big)\sigma_0.$$

$\Phi(t)$ satisfies a Riccati equation identical in form to the classical LQR gain and is interpreted as the idiosyncratic feedback gain. $\Theta(t)$ satisfies a linear equation driven by both local and macro gains and is the population coupling coefficient. $\Xi(t)$ is the tracking bias encoding the influence of $M(t)$ and $X_D$. The vanishing terminal condition $\Theta(T)=0$ reflects that the terminal cost depends only on the individual's own state, not explicitly on the population average.

3.3 Consistency Conditions and Explicit Construction of the Nash Equilibrium

Taking the conditional expectation of $p(t)=\Phi X+\Theta\bar X+\Xi$ with respect to $\mathcal F_t^0$ and comparing with $\bar p(t)=\phi\bar X+\psi$ — since both represent the same process — the uniqueness of the affine decomposition forces the Mean Field Consistency Condition:

$$\Phi(t)+\Theta(t)=\phi(t),\qquad \Xi(t)=\psi(t) \quad\Longrightarrow\quad \Theta(t)=\phi(t)-\Phi(t).$$

Explicit solution of the macro Riccati equation. Let $\mu=\tfrac{\alpha-a}2=\tfrac{-2a+b-\gamma}2$ and $\omega_2=\sqrt{\mu^2+\eta}$. The Riccati equation $\dot\phi+2\mu\phi-\phi^2+\eta=0$ has characteristic roots $\mu\pm\omega_2$, and the unique solution satisfying $\phi(T)=g$ is

$$\phi(t)=\mu+\omega_2\,\frac{(g-\mu)\cosh(\omega_2(T-t))+\omega_2\sinh(\omega_2(T-t))}{\omega_2\cosh(\omega_2(T-t))+(g-\mu)\sinh(\omega_2(T-t))}.$$

Explicit solution of the macro tracking equation. By the integrating factor method,

$$\psi(t)=-gX_D\,e^{-\int_t^T(a+\phi(s))ds}-\int_t^T \eta M(s)\, e^{-\int_t^s(a+\phi(\tau))d\tau}\,ds.$$

Explicit solution of the individual Riccati equation. Let $\omega_1=\sqrt{a^2+q+\eta}$. Proceeding as before,

$$\Phi(t)=-a+\omega_1\,\frac{(g+a)\cosh(\omega_1(T-t))+\omega_1\sinh(\omega_1(T-t))}{\omega_1\cosh(\omega_1(T-t))+(g+a)\sinh(\omega_1(T-t))}.$$

The consistency condition immediately reconstructs $\Theta(t)=\phi(t)-\Phi(t)$ and $\Xi(t)=\psi(t)$ — no additional differential equations need be solved.

Explicit Decentralized Nash Equilibrium. Assume $a>0$, $q>0$, $\eta>0$, $g>0$, and let $M\in C([0,T])$ be a deterministic continuous trajectory. Then the generalized LQ-MFG admits a unique affine decentralized Nash equilibrium, with optimal control of every representative agent given by $$u_i^*(t)=-\Phi(t)X_i(t)-\big(\phi(t)-\Phi(t)+\gamma\big)\bar X(t)-\psi(t),$$ where $\Phi(t)$ solves the individual Riccati equation, $\phi(t)$ the aggregate Riccati equation, and $\psi(t)$ the linear tracking equation. The associated state and adjoint processes satisfy the coupled FBSDE of Section 2, and the resulting aggregate trajectory satisfies the Mean Field consistency condition.

Rewriting the equilibrium controller reveals its strategic decomposition:

$$u_i^*(t)=\underbrace{-\Phi(t)\big(X_i(t)-\bar X(t)\big)}_{\text{consensus correction}}\;\underbrace{-\;(\phi(t)+\gamma)\bar X(t)}_{\text{system-wide stabilization}}\;\underbrace{-\;\psi(t)}_{\text{anticipatory tracking}}.$$
-Φ(Xᵢ-X̄) local regulation -(φ+γ)X̄ population coordination -ψ(t) exogenous tracking u*ᵢ(t)
Fig. 3 — Proposition 4.1: the equilibrium control is a superposition of three independently governed feedback layers.

Interactive Model — Simulated Population Dynamics

A finite-N Euler–Maruyama approximation of the continuum game, using the closed-form $\Phi(t)$, $\phi(t)$, $\psi(t)$ derived above and the equilibrium law $u_i^*(t)$ from Theorem 3.1. Each dot is an agent; the bright line is the empirical mean field $\bar X(t)$; the dashed line is the tracking trajectory $M(t)$.
agents Xᵢ(t) mean field X̄(t) tracking M(t)

04 Corollaries and Structural Properties of the Equilibrium

The following corollaries follow directly from Theorem 3.1 and illustrate how the generalized model reduces to well-known LQ-MFGs under appropriate parameter choices, while also isolating the independent role played by each additional modeling component.

Reduction to the Classical LQ-MFG. Suppose $\gamma=0$, $M(t)\equiv0$, $\sigma_0=0$. Then the tracking equation becomes $\dot\psi-(a+\phi)\psi=0$ with $\psi(T)=-gX_D$, the mean field becomes deterministic, and the equilibrium controller simplifies to $u_i^*(t)=-\Phi(t)X_i(t)-(\phi(t)-\Phi(t))\bar X(t)-\psi(t)$ — precisely the feedback structure of the standard LQ-MFG.
Pure Consensus Regulation. If $M(t)\equiv X_D$ is constant and $g=0$, the terminal penalty and dynamic tracking correction vanish; the decentralized control consists entirely of feedback terms involving the individual state and the aggregate population.
Absence of Common Noise. If $\sigma_0=0$, then $d\bar X=(\alpha\bar X-\bar p)dt$ is an ordinary differential equation and the aggregate trajectory is completely deterministic — consistent with the standard MFG framework in which randomness vanishes after averaging over infinitely many agents.
No Mean-Reversion. If $a=0$, the dynamics lose their Ornstein–Uhlenbeck structure: $dX_i=(b\bar X+u_i)dt+\sigma dW_i+\sigma_0dW_0$. The Riccati equations simplify to $\dot\Phi-\Phi^2+(q+\eta)=0$ and $\dot\phi+(b-\gamma)\phi-\phi^2+\eta=0$; the stabilizing influence of intrinsic mean reversion disappears, leaving the equilibrium determined solely by population interaction and tracking.
Vanishing Population Interaction. If $b=0$, the state equation reduces to $dX_i=(-aX_i+u_i)dt+\sigma dW_i+\sigma_0 dW_0$; coupling between players survives only through the quadratic consensus penalty $\tfrac q2(X_i-\bar X)^2$ — interaction may arise entirely through the objective functional even when the physical dynamics are uncoupled.
Separation of Strategic Components. The equilibrium controller admits the unique decomposition $u_i^*=u_{\text{local}}+u_{\text{macro}}+u_{\text{tracking}}$, where $u_{\text{local}}=-\Phi(X_i-\bar X)$, $u_{\text{macro}}=-(\phi+\gamma)\bar X$, and $u_{\text{tracking}}=-\psi$. Proof. Write $-\Phi X_i = -\Phi(X_i-\bar X)+\Phi\bar X$; combining the two terms multiplying $\bar X$ gives $-\Phi\bar X-(\phi-\Phi+\gamma)\bar X=-(\phi+\gamma)\bar X$, establishing the decomposition.

Because each component depends on a distinct deterministic coefficient, modifications to one modeling feature affect only its corresponding feedback channel without altering the remaining components — a modular structure that is one of the principal analytical advantages of the proposed framework.

MechanismGovernsVanishes when
Local regulation$\Phi(t)$ — individual Riccati gain$X_i \equiv \bar X$
Population coordination$\phi(t)+\gamma$ — macro gain + control–state coupling$\sigma_0=0,\ \gamma=0$ (Cor. 4.1, 4.3)
Exogenous tracking$\psi(t)$ — driven by $M(t)$, $X_D$$M\equiv X_D,\ g=0$ (Cor. 4.2)

05 Conclusions

This paper developed a generalized continuous-time Linear-Quadratic Mean Field Game incorporating four interacting mechanisms: mean-reverting dynamics, common stochastic disturbances, exogenous trajectory tracking, and a control–state interaction penalty. Using the Stochastic Maximum Principle, the decentralized optimization problem was transformed into a coupled Forward-Backward Stochastic Differential Equation. By introducing hierarchical affine representations for both the aggregate and individual adjoint processes, the original infinite-dimensional stochastic game was reduced to two nonlinear Riccati equations together with non-homogeneous linear tracking equations, enabling the complete analytical construction of the decentralized Nash equilibrium without iterative numerical procedures.

The resulting equilibrium controller naturally decomposes into three feedback layers corresponding to local consensus regulation, systemic stabilization, and exogenous trajectory tracking. Several classical LQ-MFGs were recovered as special cases by suitable parameter choices, demonstrating that the proposed formulation provides a unified framework encompassing a broad family of existing models.

Future research may extend the present analysis to multidimensional state spaces, nonlinear drift structures, heterogeneous populations, and learning-based Mean Field Games in which the exogenous trajectory is estimated online rather than prescribed a priori.

References

  1. Lasry, J.-M. and Lions, P.-L. (2007). Mean field games. Japanese Journal of Mathematics, 2(1), 229–260.
  2. Huang, M., Caines, P. E. and Malhamé, R. P. (2006). Large population stochastic dynamic games: closed-loop McKean–Vlasov systems and the Nash certainty equivalence principle. Communications in Information and Systems, 6(3), 221–252.
  3. Bensoussan, A., Frehse, J. and Yam, P. (2013). Mean Field Games and Mean Field Type Control Theory. New York: Springer.
  4. Carmona, R. and Delarue, F. (2018). Probabilistic Theory of Mean Field Games with Applications I & II. Cham: Springer.
  5. Cardaliaguet, P. (2013). Notes on Mean Field Games. Lecture notes, Université Paris-Dauphine.
  6. Carmona, R., Fouque, J.-P. and Sun, L.-H. (2015). Mean field games and systemic risk. Communications in Mathematical Sciences, 13(4), 911–933.
  7. Bensoussan, A., Sung, K. C. J., Yam, S. C. P. and Yung, S.-P. (2016). Linear-quadratic mean field games. Journal of Optimization Theory and Applications, 169(2), 496–529.
  8. Yong, J. and Zhou, X. Y. (1999). Stochastic Controls: Hamiltonian Systems and HJB Equations. New York: Springer.
  9. Vasicek, O. (1977). An equilibrium characterization of the term structure. Journal of Financial Economics, 5(2), 177–188.
  10. Carmona, R. and Zhu, X. (2016). A probabilistic approach to mean field games with major and minor players. Annals of Applied Probability, 26(3), 1535–1580.
  11. Ma, J., Protter, P. and Yong, J. (1994). Solving forward-backward stochastic differential equations explicitly — a four step scheme. Probability Theory and Related Fields, 98(3), 339–359.
  12. Peng, S. (1990). A general stochastic maximum principle for optimal control problems. SIAM Journal on Control and Optimization, 28(4), 966–979.
  13. Singh, A. (2026). A Generalized Linear-Quadratic Mean Field Game with Common Noise, Mean-Reverting Dynamics, and Exogenous Tracking. NPMAI ECOSYSTEM Research Paper, Research Mathematics Department.
NPMAI ECOSYSTEM · Research Mathematics Department
Aadarsh Singh, HOD Research Mathematics Department, NPMAI ECOSYSTEM

Aadarsh Singh

HOD, Research Mathematics Department · NPMAI ECOSYSTEM
Age 15 INMO Qualified 2026 Purple Comet — World Rank 1 AMC Perfect Score

The relationship between Aadarsh Singh and mathematics did not begin with a textbook — it began with a question. At nine years old, what drew him in was not the calculations or the formulas, but the logic underneath them: every statement had a reason, every conclusion a chain of thought behind it, and every problem a mystery waiting to be unravelled. That feeling never left him.

His father nurtured that curiosity in childhood, and his mother provided the steady support that let it grow from school interest into serious discipline. In eighth grade he was introduced to mathematical olympiads, and qualified for INMO in his very first year — followed, the next year, by a harsh recalibration: zero marks, and a clear view of how vast the world of higher mathematics really is. The year after was harder still, missing the IMOTC cutoff by five marks. He returned to training with greater discipline, working through algebra, combinatorics, and number theory — and on his third attempt, qualified for INMO again and earned a place at IMOTC. Progress in mathematics, as he puts it, is rarely linear; success tends to come after repeated failures, provided each one is learned from.

Mathematics feels like being a detective — finding patterns, watching unrelated observations fit into a complete solution.

For Aadarsh, mathematics is a way of thinking rather than a collection of problems and theorems. His immediate goal is to represent India at the International Mathematical Olympiad. Within NPMAI ECOSYSTEM, he leads the Research Mathematics Department, applying the same instinct for structure and proof to original work in stochastic control and game theory — of which this paper is the department's second publication, following his earlier work on multi-criterion congestion control and the Price of Anarchy.

INMO Qualified — 2024, 2025, 2026
NSEJS State Top 10 · NSEA State Top 8
AMC Perfect Score · AIME Qualifier
IOQM 2025 — 79/100, among India's best scores
Purple Comet — World Rank 1
Math Premier League — World Rank 1
Top 100, India — Math.Biz (2024, 2025)
NMTC Junior Qualifier — 2024, 2025
FTRE AIR-19 · FIITJEE-IITGENIUS AIR-2
TALLENTEX AIR-16 · DSSL AIR-3 · TSTMS State Rank 1

Beyond mathematics, Aadarsh co-founded a Telegram channel for math olympiad guidance and formerly administered one of India's largest math olympiad communities. Outside of research, his interests include card games and card tricks, fiction novels, cricket, chess, and thriller films.

← Back to NPMAI ECOSYSTEM