MGUP: A Momentum-Gradient Alignment Update Policy for Stochastic Optimization
NeurIPSSpotlight2025
TL;DR
Efficient optimization is essential for training large language models…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
large language model language model optimization alignment efficient
← All NeurIPS 2025 Spotlight papers · Browse the whole archive