Object Detection in Aerial Images: A Large-Scale Benchmark and Challenges

Ding, Jian; Xue, Nan; Xia, Gui-Song; Bai, Xiang; Yang, Wen; Yang, Micheal Ying; Belongie, Serge; Luo, Jiebo; Datcu, Mihai; Pelillo, Marcello; Zhang, Liangpei

doi:10.1109/TPAMI.2021.3117983

Computer Science > Computer Vision and Pattern Recognition

arXiv:2102.12219 (cs)

[Submitted on 24 Feb 2021 (v1), last revised 4 Dec 2021 (this version, v2)]

Title:Object Detection in Aerial Images: A Large-Scale Benchmark and Challenges

Authors:Jian Ding, Nan Xue, Gui-Song Xia, Xiang Bai, Wen Yang, Micheal Ying Yang, Serge Belongie, Jiebo Luo, Mihai Datcu, Marcello Pelillo, Liangpei Zhang

View PDF

Abstract:In the past decade, object detection has achieved significant progress in natural images but not in aerial images, due to the massive variations in the scale and orientation of objects caused by the bird's-eye view of aerial images. More importantly, the lack of large-scale benchmarks has become a major obstacle to the development of object detection in aerial images (ODAI). In this paper,we present a large-scale Dataset of Object deTection in Aerial images (DOTA) and comprehensive baselines for ODAI. The proposed DOTA dataset contains 1,793,658 object instances of 18 categories of oriented-bounding-box annotations collected from 11,268 aerial images. Based on this large-scale and well-annotated dataset, we build baselines covering 10 state-of-the-art algorithms with over 70 configurations, where the speed and accuracy performances of each model have been evaluated. Furthermore, we provide a code library for ODAI and build a website for evaluating different algorithms. Previous challenges run on DOTA have attracted more than 1300 teams worldwide. We believe that the expanded large-scale DOTA dataset, the extensive baselines, the code library and the challenges can facilitate the designs of robust algorithms and reproducible research on the problem of object detection in aerial images.

Comments:	Accepted to IEEE TPAMI
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
ACM classes:	I.4.8
Cite as:	arXiv:2102.12219 [cs.CV]
	(or arXiv:2102.12219v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2102.12219
Related DOI:	https://doi.org/10.1109/TPAMI.2021.3117983

Submission history

From: Gui-Song Xia [view email]
[v1] Wed, 24 Feb 2021 11:20:55 UTC (9,716 KB)
[v2] Sat, 4 Dec 2021 12:33:46 UTC (14,203 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Object Detection in Aerial Images: A Large-Scale Benchmark and Challenges

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Object Detection in Aerial Images: A Large-Scale Benchmark and Challenges

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators