Object Detection in Aerial Images: A Large-Scale Benchmark and   Challenges

Jian Ding; Nan Xue; Gui-Song Xia; Xiang Bai; Wen Yang; Micheal Ying; Yang; Serge Belongie; Jiebo Luo; Mihai Datcu; Marcello Pelillo; Liangpei; Zhang

arXiv:2102.12219·cs.CV·December 7, 2021

Object Detection in Aerial Images: A Large-Scale Benchmark and Challenges

Jian Ding, Nan Xue, Gui-Song Xia, Xiang Bai, Wen Yang, Micheal Ying, Yang, Serge Belongie, Jiebo Luo, Mihai Datcu, Marcello Pelillo, Liangpei, Zhang

PDF

2 Repos

TL;DR

This paper introduces a large-scale aerial image dataset called DOTA, along with comprehensive benchmarks and tools, to advance object detection research in aerial imagery characterized by high variability in object scale and orientation.

Contribution

The paper presents the DOTA dataset with extensive annotations, establishes multiple baseline algorithms, and provides tools and challenges to foster progress in aerial image object detection.

Findings

01

DOTA contains 1,793,658 object instances across 18 categories.

02

Baseline algorithms achieve varying accuracy and speed on DOTA.

03

Over 1300 teams participated in challenges using DOTA.

Abstract

In the past decade, object detection has achieved significant progress in natural images but not in aerial images, due to the massive variations in the scale and orientation of objects caused by the bird's-eye view of aerial images. More importantly, the lack of large-scale benchmarks has become a major obstacle to the development of object detection in aerial images (ODAI). In this paper,we present a large-scale Dataset of Object deTection in Aerial images (DOTA) and comprehensive baselines for ODAI. The proposed DOTA dataset contains 1,793,658 object instances of 18 categories of oriented-bounding-box annotations collected from 11,268 aerial images. Based on this large-scale and well-annotated dataset, we build baselines covering 10 state-of-the-art algorithms with over 70 configurations, where the speed and accuracy performances of each model have been evaluated. Furthermore, we…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.