UniTok: a Unified Tokenizer for Visual Generation and Understanding
NeurIPSSpotlight2025
TL;DR
Visual generative and understanding models typically rely on distinct tokenizers to process images, presenting a key challenge for unifying them within a single framework…
Opening excerpt from the authors’ abstract. source
Read the paper
← All NeurIPS 2025 Spotlight papers · Browse the whole archive