Understanding LLM Behaviors via Compression: Data Generation, Knowledge Acquisition and Scaling Laws
NeurIPSSpotlight2025
TL;DR
Large Language Models (LLMs) have demonstrated remarkable capabilities across numerous tasks, yet principled explanations for their underlying mechanisms and several phenomena, such as scaling laws, h…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
large language model language model scaling law llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive