To Think or Not To Think: A Study of Thinking in Rule-Based Visual Reinforcement Fine-Tuning

NeurIPSSpotlight2025

Authors
Ming Li, Jike Zhong, Shitian Zhao, Yuxiang Lai, Haoquan Zhang, Wang Bill Zhu, Kaipeng Zhang
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

This paper investigates the role of explicit thinking process in rule-based reinforcement fine-tuning (RFT) for multi-modal large language models (MLLMs)

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

large language model language model fine-tuning llm

← All NeurIPS 2025 Spotlight papers · Browse the whole archive