Frame Context Packing and Drift Prevention in Next-Frame-Prediction Video Diffusion Models

NeurIPSSpotlight2025

Authors
Lvmin Zhang, Shengqu Cai, Muyang Li, Gordon Wetzstein, Maneesh Agrawala
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

We present a neural network structure, FramePack, to train next-frame (or next-frame-section) prediction models for video generation…

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

diffusion video

← All NeurIPS 2025 Spotlight papers · Browse the whole archive