Reasoning with Sampling: Your Base Model is Smarter Than You Think

ICLROral2026

Authors
Aayush Karan, Yilun Du
Affiliation
Harvard University
Venue
ICLR 2026
Track
Oral

TL;DR

We find a training-free sampling algorithm that achieves reasoning boosts on base models comparable to those obtained by RL techniques.

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

reasoning

← All ICLR 2026 Oral papers · Browse the whole archive