MoESD: Unveil Speculative Decoding's Potential for Accelerating Sparse MoE
NeurIPSSpotlight2025
TL;DR
Large Language Models (LLMs) have achieved remarkable success across many applications, with Mixture of Experts (MoE) models demonstrating great potential…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
large language model mixture of experts language model llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive