Through the Lens of Contrast: Self-Improving Visual Reasoning in VLMs
ICLROral2026
TL;DR
Reasoning has emerged as a key capability of large language models. In linguistic tasks, this capability can be enhanced by self-improving techniques that refine reasoning paths for subsequent fine-tuning.
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
large language model language model fine-tuning reasoning