Learning from Physical Human Feedback: An Object-Centric One-Shot   Adaptation Method

Alvin Shek; Bo Ying Su; Rui Chen; Changliu Liu

arXiv:2203.04951·cs.RO·June 5, 2023

Learning from Physical Human Feedback: An Object-Centric One-Shot Adaptation Method

Alvin Shek, Bo Ying Su, Rui Chen, Changliu Liu

PDF

Open Access 1 Repo

TL;DR

This paper introduces Object Preference Adaptation (OPA), a method enabling robots to quickly adapt to human feedback in new tasks by updating object-specific preferences in a one-shot manner, without extensive retraining.

Contribution

The paper presents a novel one-shot adaptation method that updates object preferences based on minimal human feedback, improving transferability and efficiency in robot learning.

Findings

01

OPA successfully adapts to human feedback on physical robots.

02

The method requires only one human intervention for adaptation.

03

Training on synthetic data enables effective real-world application.

Abstract

For robots to be effectively deployed in novel environments and tasks, they must be able to understand the feedback expressed by humans during intervention. This can either correct undesirable behavior or indicate additional preferences. Existing methods either require repeated episodes of interactions or assume prior known reward features, which is data-inefficient and can hardly transfer to new tasks. We relax these assumptions by describing human tasks in terms of object-centric sub-tasks and interpreting physical interventions in relation to specific objects. Our method, Object Preference Adaptation (OPA), is composed of two key stages: 1) pre-training a base policy to produce a wide variety of behaviors, and 2) online-updating according to human feedback. The key to our fast, yet simple adaptation is that general interaction dynamics between agents and objects are fixed, and only…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Alvinosaur/opa
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsReinforcement Learning in Robotics · Human Pose and Action Recognition · Explainable Artificial Intelligence (XAI)

MethodsBalanced Selection