MGUP: A Momentum-Gradient Alignment Update Policy for Stochastic Optimization

NeurIPSSpotlight2025

Authors
Da Chang, Ganzhao Yuan
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

Efficient optimization is essential for training large language models…

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

large language model language model optimization alignment efficient

← All NeurIPS 2025 Spotlight papers · Browse the whole archive