Flattening Hierarchies with Policy Bootstrapping

NeurIPSSpotlight2025

Authors
John Luoyu Zhou, Jonathan Kao
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

Offline goal-conditioned reinforcement learning (GCRL) is a promising approach for pretraining generalist policies on large datasets of reward-free trajectories, akin to the self-supervised objectives…

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

reinforcement learning self-supervised pretraining dataset

← All NeurIPS 2025 Spotlight papers · Browse the whole archive