MC-Search: Evaluating and Enhancing Multimodal Agentic Search with Structured Long Reasoning Chains

ICLROral2026

Authors
Xuying Ning, Dongqi Fu, Tianxin Wei, Mengting Ai, Jiaru Zou, Ting-Wei Li, Hanghang Tong, Yada Zhu, Hendrik Hamann, Jingrui He
Affiliation
University of Illinois at Urbana-Champaign
Venue
ICLR 2026
Track
Oral

TL;DR

With the increasing demand for step-wise, cross-modal, and knowledge-grounded reasoning, multimodal large language models (MLLMs) are evolving beyond the traditional fixed retrieve-then-generate paradigm toward more sophisticated agentic multimodal retrieval-augmented generation (MM-RAG). Existing benchmarks, however, mainly focus on simplified...

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

large language model language model multimodal benchmark reasoning retrieval agent llm rag

← All ICLR 2026 Oral papers · Browse the whole archive