SATURN: SAT-based Reinforcement Learning to Unleash LLMs Reasoning
NeurIPSSpotlight2025
TL;DR
How to design reinforcement learning (RL) tasks that effectively unleash the reasoning capability of large language models (LLMs) remains an open question…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
reinforcement learning large language model language model reasoning llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive