FPSAttention: Training-Aware FP8 and Sparsity Co-Design for Fast Video Diffusion
NeurIPSSpotlight2025
TL;DR
Diffusion generative models have become the standard for producing high-quality, coherent video content, yet their slow inference speeds and high computational demands hinder practical deployment…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
generative model attention diffusion sparsity video
← All NeurIPS 2025 Spotlight papers · Browse the whole archive