GEPA: Reflective Prompt Evolution Can Outperform Reinforcement Learning
ICLROral2026
TL;DR
GEPA uses natural language reflection to optimize prompts, outperforming GRPO and MIPROv2 while needing far fewer rollouts.
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
reinforcement learning