Computer Science > Robotics

arXiv:2402.17768 (cs)

[Submitted on 27 Feb 2024 (v1), last revised 5 Jun 2024 (this version, v2)]

Title:Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning

Authors:Xiaoyu Zhang, Matthew Chang, Pranav Kumar, Saurabh Gupta

Abstract:A common failure mode for policies trained with imitation is compounding execution errors at test time. When the learned policy encounters states that are not present in the expert demonstrations, the policy fails, leading to degenerate behavior. The Dataset Aggregation, or DAgger approach to this problem simply collects more data to cover these failure states. However, in practice, this is often prohibitively expensive. In this work, we propose Diffusion Meets DAgger (DMD), a method to reap the benefits of DAgger without the cost for eye-in-hand imitation learning problems. Instead of collecting new samples to cover out-of-distribution states, DMD uses recent advances in diffusion models to synthesize these samples. This leads to robust performance from few demonstrations. We compare DMD against behavior cloning baseline across four tasks: pushing, stacking, pouring, and shirt hanging. In pushing, DMD achieves 80% success rate with as few as 8 expert demonstrations, where naive behavior cloning reaches only 20%. In stacking, DMD succeeds on average 92% of the time across 5 cups, versus 40% for BC. When pouring coffee beans, DMD transfers to another cup successfully 80% of the time. Finally, DMD attains 90% success rate for hanging shirt on a clothing rack.

Comments:	Accepted by Robotics: Science and Systems (RSS) 2024. project website with video, see this https URL
Subjects:	Robotics (cs.RO); Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
Cite as:	arXiv:2402.17768 [cs.RO]
	(or arXiv:2402.17768v2 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2402.17768

Submission history

From: Xiaoyu Zhang [view email]
[v1] Tue, 27 Feb 2024 18:59:18 UTC (10,588 KB)
[v2] Wed, 5 Jun 2024 17:33:56 UTC (42,898 KB)

Computer Science > Robotics

Title:Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Diffusion Meets DAgger: Supercharging Eye-in-hand Imitation Learning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators