LGGTJun 21

Stationary Robust Mean-Field Games under Model Mismatches

arXiv:2606.225796.9
Predicted impact top 64% in LG · last 90 daysOriginality Incremental advance
AI Analysis

For researchers in multi-agent RL, this work provides a theoretically grounded framework to handle model uncertainty in large-scale systems, though it is an incremental extension of mean-field games to robust settings.

This paper addresses model mismatches in multi-agent reinforcement learning by developing a stationary robust mean-field game framework. It proves existence of a robust equilibrium, provides the first convergent algorithm, and derives non-asymptotic error bounds, with numerical validation.

Deploying multi-agent reinforcement learning (MARL) in the real world is often limited by model mismatches between the training simulators and the true environment, which could be further amplified through strategic interactions and result in severe performance degradation upon deployment. Distributional robustness offers a principled response by optimizing policies against worst-case transition models drawn from an uncertainty set, but standard robust MARL frameworks become increasingly intractable as the number of agents grows. This paper develops an infinite-horizon, stationary mean-field game framework that incorporates distributional model uncertainty directly into the population-coupled dynamics. We establish a robust dynamic programming principle with a contractive Bellman operator and prove the existence of a stationary robust mean-field equilibrium via a fixed-point argument. We further develop the first concrete algorithm with convergence guarantees. We then connect the mean-field solution to a finite-population robust game whose ambiguity sets depend on the empirical distribution, showing that the mean-field equilibrium policy induces approximate equilibrium behavior as the population size increases. Under a contractive robust-dynamics regime, we further obtain explicit non-asymptotic error bounds. Numerical experiments further illustrate the qualitative and quantitative impact of robustness under multiple uncertainty models, validating our theoretical findings.

Foundations

The foundational work for this paper's niche, ranked by how specifically the neighbourhood builds on it — not by global fame.

Your Notes