Making Friends in the Dark: Ad Hoc Teamwork Under Partial Observability

Jo\~ao G. Ribeiroa; Cassandro Martinhoa; Alberto Sardinhaa; and; Francisco S. Melo

arXiv:2310.01439·cs.MA·October 4, 2023

Making Friends in the Dark: Ad Hoc Teamwork Under Partial Observability

Jo\~ao G. Ribeiroa, Cassandro Martinhoa, Alberto Sardinhaa, and, Francisco S. Melo

PDF

Open Access 1 Repo

TL;DR

This paper formalizes ad hoc teamwork under partial observability and introduces a model-based approach that relies solely on prior knowledge and partial observations, enabling effective collaboration without access to teammate actions or reward signals.

Contribution

It presents the first formal definition and a novel model-based method for ad hoc teamwork under partial observability with specific assumptions, advancing the field.

Findings

01

Effective in assisting unknown teammates across 70 POMDPs

02

Robust scalability to more challenging problems

03

Outperforms previous approaches in partial observability settings

Abstract

This paper introduces a formal definition of the setting of ad hoc teamwork under partial observability and proposes a first-principled model-based approach which relies only on prior knowledge and partial observations of the environment in order to perform ad hoc teamwork. We make three distinct assumptions that set it apart previous works, namely: i) the state of the environment is always partially observable, ii) the actions of the teammates are always unavailable to the ad hoc agent and iii) the ad hoc agent has no access to a reward signal which could be used to learn the task from scratch. Our results in 70 POMDPs from 11 domains show that our approach is not only effective in assisting unknown teammates in solving unknown tasks but is also robust in scaling to more challenging problems.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

jmribeiro/adhoc-teamwork-under-partial-observability
noneOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsReinforcement Learning in Robotics · Multi-Agent Systems and Negotiation · Topic Modeling

MethodsHigh-Order Consensuses