Policy Learning for Active Target Tracking over Continuous SE(3)   Trajectories

Pengzhi Yang; Shumon Koga; Arash Asgharivaskasi; Nikolay Atanasov

arXiv:2212.01498·cs.RO·May 18, 2023

Policy Learning for Active Target Tracking over Continuous SE(3) Trajectories

Pengzhi Yang, Shumon Koga, Arash Asgharivaskasi, Nikolay Atanasov

PDF

Open Access 1 Repo

TL;DR

This paper introduces a model-based policy gradient method for active target tracking with a mobile robot in continuous 3D space, optimizing sensor measurements to reduce target uncertainty.

Contribution

It presents a neural network control policy that incorporates robot pose and target distribution info, with an explicit gradient derivation for efficient optimization.

Findings

01

Effective reduction of target distribution entropy achieved

02

Neural network policy handles variable number of targets

03

Explicit gradient derivation improves training efficiency

Abstract

This paper proposes a novel model-based policy gradient algorithm for tracking dynamic targets using a mobile robot, equipped with an onboard sensor with limited field of view. The task is to obtain a continuous control policy for the mobile robot to collect sensor measurements that reduce uncertainty in the target states, measured by the target distribution entropy. We design a neural network control policy with the robot $S E (3)$ pose and the mean vector and information matrix of the joint target distribution as inputs and attention layers to handle variable numbers of targets. We also derive the gradient of the target entropy with respect to the network parameters explicitly, allowing efficient model-based policy gradient optimization.

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

existentialrobotics/rl_active_multi_target_tracking
pytorchOfficial

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.

Taxonomy

TopicsReinforcement Learning in Robotics · Gaussian Processes and Bayesian Inference · Domain Adaptation and Few-Shot Learning