Text-to-3D by Stitching a Multi-view Reconstruction Network to a Video Generator
ICLROral2026
TL;DR
Text-to-3D scene generative modelling by unifying a video generative model with a foundational 3D model via model stitching and alignment.
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
generative model alignment video 3d