ProgressLabeller: Visual Data Stream Annotation for Training Object-Centric 3D Perception

Chen, Xiaotong; Zhang, Huijie; Yu, Zeren; Lewis, Stanley; Jenkins, Odest Chadwicke

Computer Science > Robotics

arXiv:2203.00283 (cs)

[Submitted on 1 Mar 2022 (v1), last revised 1 Aug 2022 (this version, v2)]

Title:ProgressLabeller: Visual Data Stream Annotation for Training Object-Centric 3D Perception

Authors:Xiaotong Chen, Huijie Zhang, Zeren Yu, Stanley Lewis, Odest Chadwicke Jenkins

View PDF

Abstract:Visual perception tasks often require vast amounts of labelled data, including 3D poses and image space segmentation masks. The process of creating such training data sets can prove difficult or time-intensive to scale up to efficacy for general use. Consider the task of pose estimation for rigid objects. Deep neural network based approaches have shown good performance when trained on large, public datasets. However, adapting these networks for other novel objects, or fine-tuning existing models for different environments, requires significant time investment to generate newly labelled instances. Towards this end, we propose ProgressLabeller as a method for more efficiently generating large amounts of 6D pose training data from color images sequences for custom scenes in a scalable manner. ProgressLabeller is intended to also support transparent or translucent objects, for which the previous methods based on depth dense reconstruction will fail. We demonstrate the effectiveness of ProgressLabeller by rapidly create a dataset of over 1M samples with which we fine-tune a state-of-the-art pose estimation network in order to markedly improve the downstream robotic grasp success rates. ProgressLabeller is open-source at this https URL.

Comments:	IROS 2022 accepted paper; project page: this https URL
Subjects:	Robotics (cs.RO); Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2203.00283 [cs.RO]
	(or arXiv:2203.00283v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2203.00283

Submission history

From: Xiaotong Chen [view email]
[v1] Tue, 1 Mar 2022 08:04:17 UTC (34,453 KB)
[v2] Mon, 1 Aug 2022 20:09:03 UTC (34,454 KB)

Computer Science > Robotics

Title:ProgressLabeller: Visual Data Stream Annotation for Training Object-Centric 3D Perception

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:ProgressLabeller: Visual Data Stream Annotation for Training Object-Centric 3D Perception

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators