DeepDiver: Adaptive Web-Search Intensity Scaling via Reinforcement Learning
NeurIPSSpotlight2025
TL;DR
Information seeking demands iterative evidence gathering and reflective reasoning, yet large language models (LLMs) still struggle with it in open-web question answering…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
reinforcement learning large language model language model reasoning llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive