UniTok: a Unified Tokenizer for Visual Generation and Understanding

NeurIPSSpotlight2025

Authors
Chuofan Ma, Yi Jiang, Junfeng Wu, Jihan Yang, Xin Yu, Zehuan Yuan, BINGYUE PENG, XIAOJUAN QI
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

Visual generative and understanding models typically rely on distinct tokenizers to process images, presenting a key challenge for unifying them within a single framework…

Opening excerpt from the authors’ abstract. source

Read the paper

← All NeurIPS 2025 Spotlight papers · Browse the whole archive