Diffusion Model-Augmented Behavioral Cloning

Chen, Shang-Fu; Wang, Hsiang-Chun; Hsu, Ming-Hao; Lai, Chun-Mao; Sun, Shao-Hua

Computer Science > Machine Learning

arXiv:2302.13335 (cs)

[Submitted on 26 Feb 2023 (v1), last revised 3 Jun 2024 (this version, v4)]

Title:Diffusion Model-Augmented Behavioral Cloning

Authors:Shang-Fu Chen, Hsiang-Chun Wang, Ming-Hao Hsu, Chun-Mao Lai, Shao-Hua Sun

View PDF HTML (experimental)

Abstract:Imitation learning addresses the challenge of learning by observing an expert's demonstrations without access to reward signals from environments. Most existing imitation learning methods that do not require interacting with environments either model the expert distribution as the conditional probability p(a|s) (e.g., behavioral cloning, BC) or the joint probability p(s, a). Despite the simplicity of modeling the conditional probability with BC, it usually struggles with generalization. While modeling the joint probability can improve generalization performance, the inference procedure is often time-consuming, and the model can suffer from manifold overfitting. This work proposes an imitation learning framework that benefits from modeling both the conditional and joint probability of the expert distribution. Our proposed Diffusion Model-Augmented Behavioral Cloning (DBC) employs a diffusion model trained to model expert behaviors and learns a policy to optimize both the BC loss (conditional) and our proposed diffusion model loss (joint). DBC outperforms baselines in various continuous control tasks in navigation, robot arm manipulation, dexterous manipulation, and locomotion. We design additional experiments to verify the limitations of modeling either the conditional probability or the joint probability of the expert distribution, as well as compare different generative models. Ablation studies justify the effectiveness of our design choices.

Comments:	ICML 2024
Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Robotics (cs.RO)
Cite as:	arXiv:2302.13335 [cs.LG]
	(or arXiv:2302.13335v4 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2302.13335

Submission history

From: Shao-Hua Sun [view email]
[v1] Sun, 26 Feb 2023 15:40:09 UTC (6,875 KB)
[v2] Mon, 5 Jun 2023 04:39:08 UTC (7,944 KB)
[v3] Mon, 20 Nov 2023 04:52:36 UTC (6,999 KB)
[v4] Mon, 3 Jun 2024 16:17:28 UTC (6,458 KB)

Computer Science > Machine Learning

Title:Diffusion Model-Augmented Behavioral Cloning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Diffusion Model-Augmented Behavioral Cloning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators