Reinforcement Learning with Evolving Rubrics for Deep Research

ICMLOral2026

Authors
Rulin Shao, Akari Asai, Shannon Shen, Hamish Ivison, Varsha Kishore, Jingming Zhuo, Xinran Zhao, Molly Park, Samuel Finlayson, David Sontag, Tyler Murray, Sewon Min, Pradeep Dasigi, Luca Soldaini, Faeze Brahman, Scott Yih, Sherry Wu, Luke Zettlemoyer, Yoon Kim, Hannaneh Hajishirzi, Pang Wei Koh
Venue
ICML 2026
Track
Oral

TL;DR

No summary has been collected for this paper yet. Read the abstract at the authoritative source below.

Read the paper

Topics

reinforcement learning

← All ICML 2026 Oral papers · Browse the whole archive