Flattening Hierarchies with Policy Bootstrapping
NeurIPSSpotlight2025
TL;DR
Offline goal-conditioned reinforcement learning (GCRL) is a promising approach for pretraining generalist policies on large datasets of reward-free trajectories, akin to the self-supervised objectives…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
reinforcement learning self-supervised pretraining dataset
← All NeurIPS 2025 Spotlight papers · Browse the whole archive