Modular Adaptive Policy Selection for Multi-Task Imitation Learning   through Task Division

Dafni Antotsiou; Carlo Ciliberto; Tae-Kyun Kim

arXiv:2203.14855·cs.LG·May 16, 2022

Modular Adaptive Policy Selection for Multi-Task Imitation Learning through Task Division

Dafni Antotsiou, Carlo Ciliberto, Tae-Kyun Kim

PDF

Open Access 1 Repo

TL;DR

This paper presents a modular approach to multi-task imitation learning that adaptively divides tasks into shared and task-specific sub-behaviours, reducing the need for extensive demonstrations and improving performance.

Contribution

It introduces proto-policies as modules with an adaptive selector to effectively partition tasks into shared and task-specific components, enhancing multi-task learning.

Findings

01

Improves accuracy over single-task and existing multi-task methods

02

Effectively divides tasks into shared and specific sub-behaviours

03

Outperforms state-of-the-art meta-learning agents

Abstract

Deep imitation learning requires many expert demonstrations, which can be hard to obtain, especially when many tasks are involved. However, different tasks often share similarities, so learning them jointly can greatly benefit them and alleviate the need for many demonstrations. But, joint multi-task learning often suffers from negative transfer, sharing information that should be task-specific. In this work, we introduce a method to perform multi-task imitation while allowing for task-specific features. This is done by using proto-policies as modules to divide the tasks into simple sub-behaviours that can be shared. The proto-policies operate in parallel and are adaptively chosen by a selector mechanism that is jointly trained with the modules. Experiments on different sets of tasks show that our method improves upon the accuracy of single agents, task-conditioned and multi-headed…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

DaphneAntotsiou/MAPS
tfOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsMultimodal Machine Learning Applications · Domain Adaptation and Few-Shot Learning · Human Pose and Action Recognition