Reinforcement Learning for Out-of-Distribution Reasoning in LLMs: An Empirical Study on Diagnosis-Related Group Coding

NeurIPSSpotlight2025

Authors
Hanyin Wang, Zhenbang Wu, Gururaj J. Kolar, Hariprasad Reddy Korsapati, Brian Bartlett, Bryan Hull, Jimeng Sun
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

Diagnosis-Related Group (DRG) codes are essential for hospital reimbursement and operations but require labor-intensive assignment…

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

reinforcement learning reasoning llm

← All NeurIPS 2025 Spotlight papers · Browse the whole archive