MoBA: Mixture of Block Attention for Long-Context LLMs
NeurIPSSpotlight2025
TL;DR
Scaling the effective context length is essential for advancing large language models (LLMs) toward artificial general intelligence (AGI)
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
large language model language model attention llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive