Through the Lens of Contrast: Self-Improving Visual Reasoning in VLMs

ICLROral2026

Authors
Zhiyu Pan, Yizheng Wu, Jiashen Hua, Junyi Feng, Shaotian Yan, Bing Deng, Zhiguo Cao, Jieping Ye
Affiliation
Alibaba Group
Venue
ICLR 2026
Track
Oral

TL;DR

Reasoning has emerged as a key capability of large language models. In linguistic tasks, this capability can be enhanced by self-improving techniques that refine reasoning paths for subsequent fine-tuning.

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

large language model language model fine-tuning reasoning

← All ICLR 2026 Oral papers · Browse the whole archive