Transformers for Multi-Object Tracking on Point Clouds

Felicia Ruppel; Florian Faion; Claudius Gl\"aser; Klaus Dietmayer

arXiv:2205.15730·cs.CV·September 7, 2022

Transformers for Multi-Object Tracking on Point Clouds

Felicia Ruppel, Florian Faion, Claudius Gl\"aser, Klaus Dietmayer

PDF

TL;DR

This paper introduces TransMOT, a transformer-based end-to-end online tracker and detector for point cloud data, leveraging attention mechanisms to improve multi-object tracking in automotive sensor data.

Contribution

The paper presents a novel transformer architecture that unifies detection and tracking in point cloud data, using a feature-space approach and a new module for track prediction.

Findings

01

Outperforms Kalman filter-based baseline on nuScenes dataset

02

Handles sensor input at arbitrary timesteps and frame skips

03

Utilizes rich latent space for improved tracking accuracy

Abstract

We present TransMOT, a novel transformer-based end-to-end trainable online tracker and detector for point cloud data. The model utilizes a cross- and a self-attention mechanism and is applicable to lidar data in an automotive context, as well as other data types, such as radar. Both track management and the detection of new tracks are performed by the same transformer decoder module and the tracker state is encoded in feature space. With this approach, we make use of the rich latent space of the detector for tracking rather than relying on low-dimensional bounding boxes. Still, we are able to retain some of the desirable properties of traditional Kalman-filter based approaches, such as an ability to handle sensor input at arbitrary timesteps or to compensate frame skips. This is possible due to a novel module that transforms the track information from one frame to the next on…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.