Offline Guarded Safe Reinforcement Learning for Medical Treatment Optimization Strategies

NeurIPSSpotlight2025

Authors
Runze Yan, Xun Shen, Akifumi Wachi, Sebastien Gros, Anni Zhao, Xiao Hu
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

When applying offline reinforcement learning (RL) in healthcare scenarios, the out-of-distribution (OOD) issues pose significant risks, as inappropriate generalization beyond clinical expertise can re…

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

offline reinforcement learning reinforcement learning generalization optimization

← All NeurIPS 2025 Spotlight papers · Browse the whole archive