A Single Online Agent Can Efficiently Learn Mean Field Games

Chenyu Zhang; Xu Chen; Xuan Di

arXiv:2405.03718·cs.LG·July 17, 2024

A Single Online Agent Can Efficiently Learn Mean Field Games

Chenyu Zhang, Xu Chen, Xuan Di

PDF

Open Access

TL;DR

This paper introduces a novel online, model-free learning approach for a single agent to efficiently find mean field Nash equilibria in large-population systems, avoiding complex iterative methods.

Contribution

It proposes the first online single-agent, model-free algorithms for learning mean field equilibria without prior knowledge of the environment.

Findings

01

The algorithms effectively approximate fixed-point iteration.

02

Sample complexity guarantees are established.

03

Numerical experiments confirm the methods' efficiency.

Abstract

Mean field games (MFGs) are a promising framework for modeling the behavior of large-population systems. However, solving MFGs can be challenging due to the coupling of forward population evolution and backward agent dynamics. Typically, obtaining mean field Nash equilibria (MFNE) involves an iterative approach where the forward and backward processes are solved alternately, known as fixed-point iteration (FPI). This method requires fully observed population propagation and agent dynamics over the entire spatial domain, which could be impractical in some real-world scenarios. To overcome this limitation, this paper introduces a novel online single-agent model-free learning scheme, which enables a single agent to learn MFNE using online samples, without prior knowledge of the state-action space, reward function, or transition dynamics. Specifically, the agent updates its policy through…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsSports Analytics and Performance · Artificial Intelligence in Games · Time Series Analysis and Forecasting