Learning to Manipulate Object Collections Using Grounded State Representations

Wilson, Matthew; Hermans, Tucker

Computer Science > Robotics

arXiv:1909.07876 (cs)

[Submitted on 17 Sep 2019 (v1), last revised 6 Aug 2020 (this version, v3)]

Title:Learning to Manipulate Object Collections Using Grounded State Representations

Authors:Matthew Wilson, Tucker Hermans

View PDF

Abstract:We propose a method for sim-to-real robot learning which exploits simulator state information in a way that scales to many objects. We first train a pair of encoder networks to capture multi-object state information in a latent space. One of these encoders is a CNN, which enables our system to operate on RGB images in the real world; the other is a graph neural network (GNN) state encoder, which directly consumes a set of raw object poses and enables more accurate reward calculation and value estimation. Once trained, we use these encoders in a reinforcement learning algorithm to train image-based policies that can manipulate many objects. We evaluate our method on the task of pushing a collection of objects to desired tabletop regions. Compared to methods which rely only on images or use fixed-length state encodings, our method achieves higher success rates, performs well in the real world without fine tuning, and generalizes to different numbers and types of objects not seen during training.

Comments:	Accepted to Conference on Robot Learning 2019 (Oral); Video results: this https URL v3: fix abstract and appendix
Subjects:	Robotics (cs.RO); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:1909.07876 [cs.RO]
	(or arXiv:1909.07876v3 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.1909.07876

Submission history

From: Matthew Wilson [view email]
[v1] Tue, 17 Sep 2019 15:06:25 UTC (4,413 KB)
[v2] Fri, 15 Nov 2019 21:39:08 UTC (4,421 KB)
[v3] Thu, 6 Aug 2020 21:54:44 UTC (4,421 KB)

Computer Science > Robotics

Title:Learning to Manipulate Object Collections Using Grounded State Representations

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Learning to Manipulate Object Collections Using Grounded State Representations

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators