Seeing Behind Objects for 3D Multi-Object Tracking in RGB-D Sequences

Norman M\"uller; Yu-Shiang Wong; Niloy J. Mitra; Angela Dai and; Matthias Nie{\ss}ner

arXiv:2012.08197·cs.CV·December 17, 2020

Seeing Behind Objects for 3D Multi-Object Tracking in RGB-D Sequences

Norman M\"uller, Yu-Shiang Wong, Niloy J. Mitra, Angela Dai and, Matthias Nie{\ss}ner

PDF

TL;DR

This paper introduces a method for 3D multi-object tracking in RGB-D sequences that leverages complete object geometry inference to improve robustness and accuracy, especially under occlusions and appearance changes.

Contribution

It proposes jointly inferring object geometry and tracking, hallucinating unseen regions to enhance correspondence and tracking robustness in RGB-D data.

Findings

01

Achieves state-of-the-art performance on dynamic object tracking.

02

Object completion improves tracking accuracy by 6.5% in mean MOTA.

03

Robust tracking under occlusion and appearance change.

Abstract

Multi-object tracking from RGB-D video sequences is a challenging problem due to the combination of changing viewpoints, motion, and occlusions over time. We observe that having the complete geometry of objects aids in their tracking, and thus propose to jointly infer the complete geometry of objects as well as track them, for rigidly moving objects over time. Our key insight is that inferring the complete geometry of the objects significantly helps in tracking. By hallucinating unseen regions of objects, we can obtain additional correspondences between the same instance, thus providing robust tracking even under strong change of appearance. From a sequence of RGB-D frames, we detect objects in each frame and learn to predict their complete object geometry as well as a dense correspondence mapping into a canonical space. This allows us to derive 6DoF poses for the objects in each frame,…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.