MOTSynth: How Can Synthetic Data Help Pedestrian Detection and Tracking?

Matteo Fabbri; Guillem Braso; Gianluca Maugeri; Orcun Cetintas,; Riccardo Gasparini; Aljosa Osep; Simone Calderara; Laura Leal-Taixe; Rita; Cucchiara

arXiv:2108.09518·cs.CV·August 24, 2021

MOTSynth: How Can Synthetic Data Help Pedestrian Detection and Tracking?

Matteo Fabbri, Guillem Braso, Gianluca Maugeri, Orcun Cetintas,, Riccardo Gasparini, Aljosa Osep, Simone Calderara, Laura Leal-Taixe, Rita, Cucchiara

PDF

1 Repo

TL;DR

This paper introduces MOTSynth, a large synthetic dataset created with a rendering engine, demonstrating its effectiveness as a substitute for real data in pedestrian detection and tracking tasks, addressing privacy and annotation challenges.

Contribution

MOTSynth is a novel, diverse synthetic dataset that can replace real data for training in pedestrian detection and tracking, reducing privacy and annotation issues.

Findings

01

MOTSynth achieves comparable performance to real data in detection tasks.

02

Synthetic data improves privacy and reduces annotation effort.

03

Experiments validate MOTSynth's effectiveness across multiple tasks.

Abstract

Deep learning-based methods for video pedestrian detection and tracking require large volumes of training data to achieve good performance. However, data acquisition in crowded public environments raises data privacy concerns -- we are not allowed to simply record and store data without the explicit consent of all participants. Furthermore, the annotation of such data for computer vision applications usually requires a substantial amount of manual effort, especially in the video domain. Labeling instances of pedestrians in highly crowded scenarios can be challenging even for human annotators and may introduce errors in the training data. In this paper, we study how we can advance different aspects of multi-person tracking using solely synthetic data. To this end, we generate MOTSynth, a large, highly diverse synthetic dataset for object detection and tracking using a rendering game…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

dvl-tum/motsynth-baselines
pytorch

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.