Influencing Long-Term Behavior in Multiagent Reinforcement Learning

Dong-Ki Kim; Matthew Riemer; Miao Liu; Jakob N. Foerster; Michael; Everett; Chuangchuang Sun; Gerald Tesauro; Jonathan P. How

arXiv:2203.03535·cs.LG·October 18, 2022·5 cites

Influencing Long-Term Behavior in Multiagent Reinforcement Learning

Dong-Ki Kim, Matthew Riemer, Miao Liu, Jakob N. Foerster, Michael, Everett, Chuangchuang Sun, Gerald Tesauro, Jonathan P. How

PDF

Open Access 1 Repo 1 Video

TL;DR

This paper introduces a long-term, farsighted framework for multiagent reinforcement learning that considers the limiting policies of other agents, leading to improved convergence and performance in complex environments.

Contribution

It proposes a novel optimization framework that accounts for the asymptotic behavior of other agents, addressing the limitations of previous myopic approaches.

Findings

01

Outperforms state-of-the-art baselines in diverse benchmarks.

02

Effectively influences long-term policy convergence.

03

Demonstrates improved stability and scalability in multiagent settings.

Abstract

The main challenge of multiagent reinforcement learning is the difficulty of learning useful policies in the presence of other simultaneously learning agents whose changing behaviors jointly affect the environment's transition and reward dynamics. An effective approach that has recently emerged for addressing this non-stationarity is for each agent to anticipate the learning of other agents and influence the evolution of future policies towards desirable behavior for its own benefit. Unfortunately, previous approaches for achieving this suffer from myopic evaluation, considering only a finite number of policy updates. As such, these methods can only influence transient future policies rather than achieving the promise of scalable equilibrium selection approaches that influence the behavior at convergence. In this paper, we propose a principled framework for considering the limiting…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

dkkim93/further
pytorchOfficial

Videos

Influencing Long-Term Behavior in Multiagent Reinforcement Learning· slideslive

Taxonomy

TopicsReinforcement Learning in Robotics · Data Stream Mining Techniques