EvoLM: In Search of Lost Training Dynamics for Language Model Reasoning
NeurIPSOral2025
TL;DR
Modern language model (LM) training has been divided into multiple stages, making it difficult for downstream developers to evaluate the impact of design choices made at each stage. We present EvoLM, a model suite that enables systematic and transparent analysis of LMs' training dynamics across pre-training, continued pre-training, supervised fine-…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
language model pre-training reasoning