Adaptive t-Momentum-based Optimization for Unknown Ratio of Outliers in Amateur Data in Imitation Learning

Ilboudo, Wendyam Eric Lionel; Kobayashi, Taisuke; Sugimoto, Kenji

Computer Science > Machine Learning

arXiv:2108.00625 (cs)

[Submitted on 2 Aug 2021]

Title:Adaptive t-Momentum-based Optimization for Unknown Ratio of Outliers in Amateur Data in Imitation Learning

Authors:Wendyam Eric Lionel Ilboudo, Taisuke Kobayashi, Kenji Sugimoto

View PDF

Abstract:Behavioral cloning (BC) bears a high potential for safe and direct transfer of human skills to robots. However, demonstrations performed by human operators often contain noise or imperfect behaviors that can affect the efficiency of the imitator if left unchecked. In order to allow the imitators to effectively learn from imperfect demonstrations, we propose to employ the robust t-momentum optimization algorithm. This algorithm builds on the Student's t-distribution in order to deal with heavy-tailed data and reduce the effect of outlying observations. We extend the t-momentum algorithm to allow for an adaptive and automatic robustness and show empirically how the algorithm can be used to produce robust BC imitators against datasets with unknown heaviness. Indeed, the imitators trained with the t-momentum-based Adam optimizers displayed robustness to imperfect demonstrations on two different manipulation tasks with different robots and revealed the capability to take advantage of the additional data while reducing the adverse effect of non-optimal behaviors.

Comments:	7 pages, Accepted in IROS 2021, See video pitch with main result on this https URL
Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2108.00625 [cs.LG]
	(or arXiv:2108.00625v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2108.00625

Submission history

From: Wendyam Eric Lionel Ilboudo [view email]
[v1] Mon, 2 Aug 2021 04:30:41 UTC (10,344 KB)

Computer Science > Machine Learning

Title:Adaptive t-Momentum-based Optimization for Unknown Ratio of Outliers in Amateur Data in Imitation Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Adaptive t-Momentum-based Optimization for Unknown Ratio of Outliers in Amateur Data in Imitation Learning

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators