LLM-Explorer: A Plug-in Reinforcement Learning Policy Exploration Enhancement Driven by Large Language Models

NeurIPSSpotlight2025

Authors
Qianyue Hao, Yiwen Song, Qingmin Liao, Jian Yuan, Yong Li
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

Policy exploration is critical in reinforcement learning (RL), where existing approaches include $\epsilon$-greedy, Gaussian process, etc…

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

reinforcement learning large language model gaussian process language model exploration llm

← All NeurIPS 2025 Spotlight papers · Browse the whole archive