Learning by Watching

Jimuyang Zhang; Eshed Ohn-Bar

arXiv:2106.05966·cs.CV·June 11, 2021

Learning by Watching

Jimuyang Zhang, Eshed Ohn-Bar

PDF

TL;DR

The paper introduces Learning by Watching (LbW), a novel framework that enables autonomous driving agents to learn from indirect observations of other vehicles, improving data efficiency and robustness without full state or action knowledge.

Contribution

LbW allows learning driving policies by observing other vehicles' behaviors through viewpoint transformation and action inference, reducing data requirements and enhancing adaptability.

Findings

01

Achieves 92% success rate with 30 minutes of data on CARLA benchmark.

02

Attains 82% success rate with only 10 minutes of data.

03

Enables robust driving policy learning without full state or action access.

Abstract

When in a new situation or geographical location, human drivers have an extraordinary ability to watch others and learn maneuvers that they themselves may have never performed. In contrast, existing techniques for learning to drive preclude such a possibility as they assume direct access to an instrumented ego-vehicle with fully known observations and expert driver actions. However, such measurements cannot be directly accessed for the non-ego vehicles when learning by watching others. Therefore, in an application where data is regarded as a highly valuable asset, current approaches completely discard the vast portion of the training data that can be potentially obtained through indirect observation of surrounding vehicles. Motivated by this key insight, we propose the Learning by Watching (LbW) framework which enables learning a driving policy without requiring full knowledge of…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

MethodsEntropy Regularization · Proximal Policy Optimization · CARLA: An Open Urban Driving Simulator