Co-Evolving LLM Coder and Unit Tester via Reinforcement Learning
NeurIPSSpotlight2025
TL;DR
Mathematical reasoning in large language models has been successfully incentivized through reinforcement learning with verifiable rewards, leading to improved one-shot precision…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
reinforcement learning large language model language model reasoning llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive