Exploratory Diffusion Model for Unsupervised Reinforcement Learning

ICLROral2026

Authors
Chengyang Ying, Huayu Chen, Xinning Zhou, Zhongkai Hao, Hang Su, Jun Zhu
Affiliation
Tsinghua University, Tsinghua University
Venue
ICLR 2026
Track
Oral

TL;DR

We propose Exploratory Diffusion Model (ExDM), boosting unsupervised exploration and few-shot fine-tuning by diffusion models.

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

reinforcement learning exploration fine-tuning diffusion few-shot

← All ICLR 2026 Oral papers · Browse the whole archive