To Think or Not To Think: A Study of Thinking in Rule-Based Visual Reinforcement Fine-Tuning
NeurIPSSpotlight2025
TL;DR
This paper investigates the role of explicit thinking process in rule-based reinforcement fine-tuning (RFT) for multi-modal large language models (MLLMs)
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
large language model language model fine-tuning llm
← All NeurIPS 2025 Spotlight papers · Browse the whole archive