BOP Challenge 2024 on Model-Based and Model-Free 6D Object Pose Estimation

Nguyen, Van Nguyen; Tyree, Stephen; Guo, Andrew; Fourmy, Mederic; Gouda, Anas; Lee, Taeyeop; Moon, Sungphill; Son, Hyeontae; Ranftl, Lukas; Tremblay, Jonathan; Brachmann, Eric; Drost, Bertram; Lepetit, Vincent; Rother, Carsten; Birchfield, Stan; Matas, Jiri; Labbe, Yann; Sundermeyer, Martin; Hodan, Tomas

Computer Science > Computer Vision and Pattern Recognition

arXiv:2504.02812 (cs)

[Submitted on 3 Apr 2025 (v1), last revised 23 Apr 2025 (this version, v4)]

Title:BOP Challenge 2024 on Model-Based and Model-Free 6D Object Pose Estimation

Authors:Van Nguyen Nguyen, Stephen Tyree, Andrew Guo, Mederic Fourmy, Anas Gouda, Taeyeop Lee, Sungphill Moon, Hyeontae Son, Lukas Ranftl, Jonathan Tremblay, Eric Brachmann, Bertram Drost, Vincent Lepetit, Carsten Rother, Stan Birchfield, Jiri Matas, Yann Labbe, Martin Sundermeyer, Tomas Hodan

View PDF HTML (experimental)

Abstract:We present the evaluation methodology, datasets and results of the BOP Challenge 2024, the 6th in a series of public competitions organized to capture the state of the art in 6D object pose estimation and related tasks. In 2024, our goal was to transition BOP from lab-like setups to real-world scenarios. First, we introduced new model-free tasks, where no 3D object models are available and methods need to onboard objects just from provided reference videos. Second, we defined a new, more practical 6D object detection task where identities of objects visible in a test image are not provided as input. Third, we introduced new BOP-H3 datasets recorded with high-resolution sensors and AR/VR headsets, closely resembling real-world scenarios. BOP-H3 include 3D models and onboarding videos to support both model-based and model-free tasks. Participants competed on seven challenge tracks. Notably, the best 2024 method for model-based 6D localization of unseen objects (FreeZeV2.1) achieves 22% higher accuracy on BOP-Classic-Core than the best 2023 method (GenFlow), and is only 4% behind the best 2023 method for seen objects (GPose2023) although being significantly slower (24.9 vs 2.7s per image). A more practical 2024 method for this task is Co-op which takes only 0.8s per image and is 13% more accurate than GenFlow. Methods have similar rankings on 6D detection as on 6D localization but higher run time. On model-based 2D detection of unseen objects, the best 2024 method (MUSE) achieves 21--29% relative improvement compared to the best 2023 method (CNOS). However, the 2D detection accuracy for unseen objects is still -35% behind the accuracy for seen objects (GDet2023), and the 2D detection stage is consequently the main bottleneck of existing pipelines for 6D localization/detection of unseen objects. The online evaluation system stays open and is available at this http URL

Comments:	arXiv admin note: text overlap with arXiv:2403.09799
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2504.02812 [cs.CV]
	(or arXiv:2504.02812v4 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2504.02812

Submission history

From: Van Nguyen Nguyen [view email]
[v1] Thu, 3 Apr 2025 17:55:19 UTC (27,237 KB)
[v2] Thu, 17 Apr 2025 18:06:29 UTC (27,217 KB)
[v3] Tue, 22 Apr 2025 10:16:03 UTC (27,152 KB)
[v4] Wed, 23 Apr 2025 11:37:45 UTC (17,713 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:BOP Challenge 2024 on Model-Based and Model-Free 6D Object Pose Estimation

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:BOP Challenge 2024 on Model-Based and Model-Free 6D Object Pose Estimation

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators