SpotlightTodAI › ICML 2026 Oral
ICML 2026 Oral Papers
ICML Oral 2026
All 168 papers accepted as Oral at ICML 2026 — the top slice of the accepted programme. Every title links to its own page with authors, affiliation and sources.
Other venues
NeurIPS 2025 Spotlight 687
ICLR 2026 Oral 224
NeurIPS 2025 Oral 77
ICML 2026 Spotlight 6
All 168 papers
Benchmarking at the Edge of Comprehension Samuele Marro, Jialin Yu, Emanuele La Malfa et al.
Asymmetric Perturbation in Solving Bilinear Saddle-Point Optimization Kenshi Abe, Mitsuki Sakamoto, Kaito Ariu et al.
Position: Don't Just "Fix it in Post'': A Science of AI Must Study Learning Dynamics Stella Biderman, Mohammad Aflah Khan, Fatemehsadat Mireshghallah et al.
dnaHNet: A Scalable and Hierarchical Foundation Model for Genomic Sequence Learning Arnav Shah, Junzhe Li, Parsa Idehpour et al.
Are VLMs Seeing or Just Saying? Uncovering the Illusion of Visual Re-examination Chufan Shi, Cheng Yang, Yaokang Wu et al.
Do We Need Adam? Surprisingly Strong and Sparse Reinforcement Learning with SGD in LLMs Sagnik Mukherjee, Lifan Yuan, Pavan Jayasinha et al.
DiScoFormer: Plug-In Density and Score Estimation with Transformers Vasily Ilin, Peter Sushko, Ranjay Krishna
CLEAR: Context-Aware Learning with End-to-End Mask-Free Inference for Adaptive Subtitle Removal Qingdong He, Chaoyi Wang, Peng TANG et al.
A Systematic Study of Behavioral Cloning for Scientific Data Annotation Ishaan Singh Chandok, Core Francisco Park
Mixtures Closest To A Given Measure: A Semidefinite Programming Approach Srećko Ðurašinović, Jean B Lasserre, Victor Magron
FLIP2: Expanding Protein Fitness Landscape Benchmarks for Real-World Machine Learning Applications Kieran Didi, Sarah Alamdari, Alex Lu et al.
daVinci-Dev: Agent-native Mid-training for Software Engineering Ji Zeng, Dayuan Fu, Tiantian Mi et al.
Learning Unmasking Policies for Diffusion Language Models Metod Jazbec, Theo X. Olausson, Louis Béthune et al.
LASER: Learning Active Sensing for Continuum Field Reconstruction Huayu Deng, Jinghui Zhong, Xiangming Zhu et al.
Protein Autoregressive Modeling via Multiscale Structure Generation Yanru Qu, Cheng-Yen Hsieh, Zaixiang Zheng et al.
Multimodal Nested Learning for Decoupled and Coordinated Optimization Yanglin Feng, Yang Qin, Dezhong Peng et al.
AI Engram: In Search of Memory Traces in Artificial Intelligence Jea Kwon, Dong-Kyum Kim, Jiwon Kim et al.
On the Convergence Rate of LoRA Gradient Descent Siqiao Mu, Diego Klabjan
Motion Attribution for Video Generation Xindi Wu, Despoina Paschalidou, Jun Gao et al.
Less is Enough: Synthesizing Diverse Data in Feature Space of LLMs Zhongzhi Li, Xuansheng Wu, Yijiang Li et al.
Strategic Navigation or Stochastic Search? How Agents and Humans Reason Over Document Collections Lukasz Borchmann, Jordy Van Landeghem, Michał Turski et al.
Protein Fold Classification at Scale: Benchmarking and Pretraining Dexiong Chen, Andrei Manolache, Mathias Niepert et al.
VenusBench-Mobile: A Challenging and User-Centric Benchmark for Mobile GUI Agents with Capability Diagnostics Yichen Gong, Zhuohan Cai, Sunhao Dai et al.
OPUS: Towards Efficient and Principled Data Selection in Large Language Model Pre-training in Every Iteration Shaobo Wang, Xuan Ouyang, Tianyi Xu et al.
Guaranteed Optimal Compositional Explanations for Neurons Biagio La Rosa, Leilani Gilpin
Revenue Efficiency of Correlated Equilibria in First Price Auctions Anders Bo Ipsen, Stratis Skoulakis · Aarhus U
Riemannian Metric Matching for Scalable Geometric Modeling of Distributions Jacob Bamberger, Adam Gosztolai, Pierre Vandergheynst et al.
PhotoAgent: Exploratory Visual Aesthetic Planning with Large Vision Models Mingde Yao, Zhiyuan You, King-Man Tam et al.
Position: Anthropomorphic Misalignment Research Needs Stronger Evidence Vansh Gupta, Peter Nutter, Samuel Stante et al.
Diffract: Spectral View of LLM Domain Adaptation Nikita Borodin, Maria Krylova, Artem Zabolotnyi et al.
Learning Human-Robot Collaboration via Heterogeneous-Agent Lyapunov Policy Optimization Hao Zhang, Yaru Niu, Yikai Wang et al.
From Feasible to Practical: Pareto-Optimal Synthesis Planning Friedrich Hastedt, Dongda Zhang, Antonio Del rio chanona
Controlled LLM Training on Spectral Sphere Tian Xie, Haoming Luo, Haoyu Tang et al.
Don't Force the Fit: Bounded Log-Likelihood Loss for Enhanced Reasoning in Large Language Models Feng Zhao, Hong Zhang, Yu Yang et al.
Chebyshev Policies and the Mountain Car Problem: Reinforcement Learning for Low-dimensional Control Tasks Stefan Huber, Hannes Unger, Georg Schäfer et al.
Monitoring Monitorability Melody Guan, Miles Wang, Micah Carroll et al.
Optimal and Scalable MAPF via Multi-Marginal Optimal Transport and Schrödinger Bridges Usman A Khan, Joseph Durham
ReViT: Rotational-equivariant Vision Transformers for Neural PDE Solvers Hao Wei, Björn List, Nils Thuerey
Evaluating Robustness of Reasoning Models on Parameterized Logical Problems Naïm Es-sebbani, Esteban Marquer, Yakoub Salhi et al.
Detecting the Semantic Fixed Point: A Geometric Framework for Efficient Inference Jiawei Gu, Ziyue Qiao, Xiao Luo
Maximum Likelihood Reinforcement Learning Fahim Tajwar, Guanning Zeng, Yueer Zhou et al.
Foundations of Equivariant Deep Learning: Unifying Graph and Sheaf Neural Networks Yoshihiro Maruyama
Rex: A Family of Reversible Exponential (Stochastic) Runge-Kutta Solvers Zander Blasingame, Chen Liu
ECHO: Elastic Speculative Decoding with Sparse Gating for High-Concurrency Scenarios Xinyi Hu, Yuhao Shen, Zhang Baolin et al.
The Obfuscation Atlas: Mapping Where Honesty Emerges in RLVR with Deception Probes Mohammad Taufeeque, Stefan Heimersheim, Adam Gleave et al.
On Minimum Depth and Width of Floating-Point Neural Networks for Representing Floating-Point Functions Sejun Park, Yeachan Park, Geonho Hwang
RoboMME: Benchmarking and Understanding Memory for Robotic Generalist Policies Yinpei Dai, Hongze Fu, Jayjun Lee et al.
Minimax Optimal Strategy for Delayed Observations in Online Reinforcement Learning Harin Lee, Kevin Jamieson
TG-RAG: A Retrieval-Augmented Framework for Reasoning Guidance in Specialized Domains Liang Su, Mingyang Zhang, Yun Xiong et al.
The Expressivity Limits of Transformers Maxime Meyer, Mario Michelessa, Caroline Chaux et al.
Optimal Decision-Making Based on Prediction Sets Tao Wang, Edgar Dobriban · UPenn
Towards Sub-second Biological Foundation Model Infrastructure: A Quantized Consistency Diffusion Framework for Molecular Docking Kexin Zhang, Weichen Qin, Yue Teng et al.
Understanding Reasoning Collapse in LLM Agent Reinforcement Learning Zihan (Zenus) Wang, Chi Gui, Xing Jin et al.
VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models Woojin Kim, Sieun Hyeon, Jusang Oh et al.
Unsupervised Partner Design Enables Robust Ad-hoc Teamwork Constantin Ruhdorfer, Matteo Bortoletto, Victor Oei et al.
MuonSSM: Orthogonalizing State Space Models for Sequence Modeling Thai Khanh Nguyen, Ngoc Bich Uyen Vo, Thieu Vo et al.
Mitigating Reward Hacking in RLHF via Bayesian Non-negative Reward Modeling Zhibin Duan, Guowei Rong, Zhuo Li et al.
Any-Order GPT as Masked Diffusion Model: Decoupling Formulation and Architecture Shuchen Xue, Tianyu Xie, Tianyang Hu et al.
Position: Irresponsible AI: big tech’s influence on AI research and associated impacts Alex Hernandez-Garcia, Alexandra Volokhova, Ezekiel Williams et al.
Bad Seeing or Bad Thinking? Rewarding Perception for Multimodal Reasoning WANG, Qixin Xu, Changpeng Wang et al.
Position: Stop Automating Peer Review Without Rigorous Evaluation Joachim Baumann, Jiaxin Pei, Sanmi Koyejo et al.
Jailbreak Foundry: From Papers to Runnable Attacks for Reproducible Benchmarking Zhicheng Fang, Jingjie Zheng, Chenxu Fu et al.
Exact Functional ANOVA Decomposition for Categorical Inputs Baptiste Ferrere, Nicolas Bousquet, Gamboa Fabrice et al.
Quantifying Frontier LLM Capabilities for Container Sandbox Escape Rahul Marchand, Art Cathain, Jerome Wynne et al.
Joint Learning in the Gaussian Single Index Model Loucas Pillaud-Vivien, Adrien Schertzer
Position: The AI Imperative: Scaling High-Quality Peer Review in Machine Learning Qiyao Wei, Samuel Holt, Jing Yang et al.
DroneDINO: Towards Heterogeneous Routed Mixture of Experts for Drone-based Unified Object Detection Rui Chen, Dongdong Li, Yan Fan et al.
Error Propagation Mechanisms and Compensation Strategies for Quantized Diffusion Models Songwei Liu, Chao Zeng, Chenqian Yan et al.
Activation Oracles: Training and Evaluating LLMs as General-Purpose Activation Explainers Adam Karvonen, James Chua, Clément Dumas et al.
Reinforcement Learning with Evolving Rubrics for Deep Research Rulin Shao, Akari Asai, Shannon Shen et al.
Simultaneous Speech-to-Speech Translation Without Aligned Data Tom Labiausse, Romain Fabre, Yannick Estève et al.
Incentivizing Truthfulness and Collaborative Fairness in Bayesian Learning Rachael Hwee Ling Sim, Jue Fan, Xiao Tian et al.
SVRG and Beyond via Posterior Correction Nico Daheim, Thomas Moellenhoff, James Ming Liang Ang et al. · TU Darmstadt / hessian.AI
Lottery Prior: Randomized Neural Compression for Zero-Shot Inverse Problems Haotian Wu, Di You, Pier Luigi Dragotti et al.
Large Language Models Develop Novel Social Biases Through Adaptive Exploration Addison J. Wu, Ryan Liu, Xuechunzi Bai et al.
High-accuracy and dimension-free sampling with diffusions Khashayar Gatmiry, Sitan Chen, Adil Salim
Robust Harmful Features Under Jailbreak Attacks: Mechanistic Evidence from Attention Head Specialization in Large Language Models Yanchen Yin, Dongqi Han, Linghui Li
The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models Zanlin Ni, Shenzhi Wang, Yang Yue et al.
Is Your LLM Overcharging You? Tokenization, Transparency, and Incentives Ander Artola Velasco, Stratis Tsirtsis, Nastaran Okati et al.
Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units Jianhui Chen, Yuzhang Luo, Liangming Pan
Video-Based Optimal Transport for Feedback-Efficient Offline Preference-Based Reinforcement Learning Minh-Tung Luu, Hwanhee Kim, Younghwan Lee et al.
Scalable Event Cloud Network for Event-based Classification Hongwei Ren, Fei Ma, Xiaopeng LIN et al.
Towards Fair Sequential Decision-Making: A Causal Decomposition Approach Jiajun Chen, Jin Tian, Chris Quinn
When the Prompt Becomes Visual: Vision-Centric Jailbreak Attacks for Large Image Editing Models Jiacheng Hou, Yining Sun, Ruochong Jin et al.
A Recursive Decomposition Framework for Causal Structure Learning in the Presence of Latent Variables Zheng Li, Feng Xie, Shenglan Nie et al.
ConFlux: Multivariate Time Series in Flux, One Unified Forecast in Confluence Shiyu Wang, Yuchen Fang, Juntong Ni et al.
CAT-Q: Cost-efficient and Accurate Ternary Quantization for LLMs Shigeng Wang, Chao Li, Yangyuxuan Kang et al.
CoEvol-NO: State and Coordinate Co-Evolution with an Error-Driven Predictor-Corrector Paradigm for Neural Operator Transformer Jianqiao Zeng, Ruocheng Wang, Yanzhi Liu et al.
DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory (and Its Loss' Convexity is Dispensable) Wenxuan Zhou, Shujian Zhang, brice magdalou et al.
Position: The Alignment Community is Unintentionally Building a Censor’s Toolkit Sarah Ball, Phil Hackemann
From Abstraction to Instantiation: Learning Behavioral Representation for Vision-Language-Action Model Bing Hu, Zaijing Li, Rui Shao et al.
Information Flow Reveals When to Trust Language Models Rui Xu, Yi Chen, Jiujiu Chen et al.
From Text to Forecasts: Bridging Modality Gap with Temporal Evolution Semantic Space Lehui Li, Yuyao Wang, Jisheng Yan et al.
High-accuracy sampling for diffusion models and log-concave distributions Fan Chen, Sinho Chewi, Constantinos Daskalakis et al.
From Pixels to Tokens: A Systematic Study of Latent Action Supervision for Vision-Language-Action Models Yihan Lin, Haoyang Li, Yang Li et al.
POET-X: Memory-efficient LLM Training by Scaling Orthogonal Transformation Zeju Qiu, Lixin LIU, Adrian Weller et al.
DISCO: Mitigating Bias in Deep Learning with Conditional Distance Correlation Emre Kavak, Tom Nuno Wolf, Christian Wachinger
Geometric Flow Grounding: A Unified Manifold Decoupling Framework for Dynamics Discovery and Verification Chang Yu, Yuxuan Luo, Yixuan Du et al.
Pretrained Vision-Language-Action Models are Surprisingly Resistant to Forgetting in Continual Learning Huihan Liu, Changyeon Kim, Bo Liu et al.
Solving Time-Dependent Differential Equations with Physical Dynamical Systems Chuan Liu, Yijie Chen, Ruibing Song et al.
Modeling Hierarchical Thinking in Large Reasoning Models G M Shahariar, Erfan Shayegani, Ali Nazari et al.
Rational Transductors Mehryar Mohri
Markov Chain Monte Carlo without Evaluating the Target: an Auxiliary Variable Approach Wei Yuan, Guanyang Wang
Disentangling Latent Risk Pathways via Bayesian Hypergraph Inference Shengxian Ding, Haonan Gao, Pangpang Liu et al.
ReQAT: Achieving Full-Precision Reasoning Accuracy with 4-bit Floating-Point Quantization-Aware Training Janghwan Lee, Sihwa Lee, Jinseok Kim et al.
Path-dependent Discrete Amortized Inference Tiago Silva, Esmeralda S. Whitammer, Salem Lahlou
To Grok Grokking: Provable Grokking in Ridge Regression Mingyue Xu, Gal Vardi, Itay Safran
Training-Free Bayesian Filtering with Generative Emulators Thomas Savary, François Rozet, Gilles Louppe
XR-1: Towards Versatile Vision-Language-Action Models via Learning Unified Vision-Motion Representations Shichao Fan, Kun Wu, Zhengping Che et al.
Skip a Layer or Loop It? Learning Program-of-Layers in LLMs Ziyue Li, Yang Li, Tianyi Zhou
On the Identifiability of Poisson Branching Structural Causal Model Under Latent Confounding Jie Qiao, Zihuai Zeng, Ruichu Cai et al.
Reward-free Alignment for Conflicting Objectives Peter Chen, Xiaopeng Li, Xi Chen et al.
Mind Your Margin and Boundary: Are Your Distilled Datasets Truly Robust? Muquan Li, Yingyi Ma, Yihong Huang et al.
Position: There are futures that benchmark-driven AI cannot see Sobhan Lotfi, Ava Iranmanesh, Lachin Naghashyar et al.
FlatLand: Personalized Graph Federated Learning via Tailored Lorentz Space Jiahong Liu, Ram Samarth B B, Xinyu Fu et al.
Necessary Conditions for Compositional Generalization of Embedding Models Arnas Uselis, Andrea Dittadi, Seong Joon Oh
3ViewSense: Spatial and Mental Perspective Reasoning from Orthographic Views in Vision-Language Models Shaoxiong Zhan, Yanlin Lai, Zheng Liu et al.
Midtraining Bridges Pretraining and Posttraining Distributions Emmy Liu, Graham Neubig, Chenyan Xiong
Distributional Inverse Reinforcement Learning Feiyang Wu, Ye Zhao, Anqi Wu
ThreadWeaver: Adaptive Threading for Efficient Parallel Reasoning in Language Models Long (Tony) Lian, Sida Wang, Felix Juefei-Xu et al.
On the Difficulty of Learning a Meta-network for Training Data Selection Zilin Du, Junqi Zhao, Albert Boyang Li
On Computation and Reinforcement Learning Raj Ghugare, Michał Bortkiewicz, Alicja Ziarko et al.
Non-Euclidean Gradient Descent Operates at the Edge of Stability Rustem Islamov, Michael Crawshaw, Jeremy Cohen et al.
CausalGame: Benchmarking Causal Thinking of LLM Agents in Games Zhenhao Chen, Yongqiang Chen, Chenxi Liu et al.
MV-FGAD: Towards Efficient and Effective Federated Graph Anomaly Detection via Multi-view Learning Junyi Yan, KE LIANG, Hao Yu et al.
Agent0-VL: Exploring Self-Evolving Agent for Tool-Integrated Vision-Language Reasoning Jiaqi Liu, Kaiwen Xiong, Peng Xia et al.
PRISM: Gauge-Invariant Tangent-Space Differentially Private LoRA Shihao Wang, Xueru Zhang
TokSuite: Measuring the Impact of Tokenizer Choice on Language Model Behavior Gül Sena Altıntaş, Malikeh Ehghaghi, Brian Lester et al.
The Signal is in the Steps: Local Scoring for Reasoning Data Selection Hoang Anh Just, Myeongseob Ko, Ruoxi Jia
Second-Order Smooth Planning with Optimal-Transport Bellman Smoothing Tuan Dam
Characterizing, Evaluating, and Optimizing Complex Reasoning Haoran Zhang, Yafu Li, Zhi Wang et al.
PhenoBrain: Phenotype-Conditioned Long-Range Communication for Multi-Modal Brain Network Analysis Lingyuan Meng, KE LIANG, Hao Li et al.
Holi-Spatial: Evolving Video Streams into Holistic 3D Spatial Intelligence Yuanyuan Gao, Hao Li, Yifei Liu et al.
Robust Contextual Optimization with Missing Covariates Qingyuan Xu, Ruiwei Jiang
Towards Hierarchy–Uniformity Equilibrium: Recovering Semantic Depth in Hypergraph Contrastive Learning Ruiting Zhao, Ming Li, Lixin Cui et al.
Transforming Weather Data from Pixel to Latent Space Sijie Zhao, Feng Liu, Xueliang Zhang et al.
SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models jing wu, Jianhua Wu, Jiayi Guan et al.
Stabilizing the Q-Gradient Field for Policy Smoothness in Actor-Critic Methods Jeong Woon Lee, Kyoleen Kwak, Daeho Kim et al.
Rare Event Analysis of Large Language Models Jake McAllister Dorman, Edward Gillman, Dominic C Rose et al.
WeDLM: Reconciling Diffusion Language Models with Standard Causal Attention for Fast Inference Aiwei Liu, Minghua He, Shaoxun Zeng et al.
A Random Matrix Perspective on the Consistency of Diffusion Models Binxu Wang, Jacob A Zavatone-Veth, Cengiz Pehlevan
Expressivity-Efficiency Tradeoffs for Hybrid Sequence Models John Cooper, Mingchen Ma, Ilias Diakonikolas et al.
$\tau^2$-Bench: Evaluating Conversational Agents in a Dual-Control Environment Victor Barres, Honghua Dong, Soham Ray et al.
DOUBT: Decoupled Object-level Understanding and Bridging via vMF-based Trustworthiness for Hallucination Detection in MLLMs Kaiqi Chen, Yang Qin, Changhao He et al.
Faster Activation Functions at the Edge for Post-Training Speedups Anton Lydike, Jun Bi, Jackson Woodruff
Position: AI Should Facilitate Democratic Deliberation at Scale José Ramón Enríquez, Jiaxin Pei, Alex Pentland
Position: AI/ML Deepfake Research is Misaligned with AI Generated Non-Consensual Intimate Imagery (AIG-NCII) Qiwei Li, Wells Lucas Santo, Sarita Schoenebeck et al.
Equilibrium Pricing in Oligopolistic Data Markets Bhaskar Ray Chaudhury, Jugal Garg, Eklavya Sharma et al.
FlashSinkhorn: IO-Aware Entropic Optimal Transport on GPU Felix X.-F. Ye, Xingjie Li, An Yu et al.
Characterizing Agents in Production Melissa Pan, Negar Arabzadeh, Riccardo Cogo et al.
Learning to Theorize the World from Observation Doojin Baek, Gyubin Lee, Junyeob Baek et al.
How much can language models memorize? John Morris, Chawin Sitawarin, Narine Kokhlikyan et al.
Equivalence of Context and Parameter Updates in Modern Transformer Blocks Adrian Goldwaser, Michael Munn, Xavi Gonzalvo et al.
GoodDiffusion: Proactive Copyright Protection for Diffusion Generative Models via Learnable Sample-specific Signatures Shixi Qin, zhiyong yang, Shilong Bao et al.
Focus and Dilution: The Multi-stage Learning Process of Attention Zheng-An Chen, Pengxiao Lin, Zhi-Qin John Xu et al.
FlashSketch: Sketch-Kernel Co-Design for Fast Sparse Sketching on GPUs Rajat Vadiraj Dwaraknath, Sungyoon Kim, Mert Pilanci
Nash Equilibria in Games with Playerwise Concave Coupling Constraints: Existence and Computation Philip Jordan, Maryam Kamgarpour
CVE-Factory: Scaling Expert-Level Agentic Tasks for Code Security Vulnerability Xianzhen Luo, Jingyuan Zhang, Shiqi Zhou et al.
Orthogonal Concept Erasure for Diffusion Models Yuhao Sun, Lingyun Yu, Hao-Xiang Xu et al.
Prescriptive Scaling Reveals the Evolution of Language Model Capabilities Hanlin Zhang, Jikai Jin, Vasilis Syrgkanis et al.
On the Limits of LLM Adaptability: Impact of LLM Pre-Training on Annotation Task Performance Etienne Casanova, Rafal Kocielnik, R. Michael Alvarez
Privacy-Aware Video Anomaly Detection: Guided Orthogonal Projection and a Comprehensive Evaluation Framework Wenxiang Diao, Lei Wang, Andrew Busch et al.
Procedural Pretraining: Warming Up Language Models with Abstract Data Liangze Jiang, Zachary Shinnick, Anton Hengel et al.
Towards Long-Horizon Interpretability: Efficient and Faithful Multi-Token Attribution for Reasoning LLMs Wenbo Pan, Zhichao Liu, Xianlong Wang et al.
Which Algorithms Can Graph Neural Networks Learn? Solveig Wittig, Antonis Vasileiou, Robert R. Nerem et al.
What Preferences Can—and Cannot—Predict in Multi-Agent Online Learning Omar Abbadi, Rida Laraki, Panayotis Mertikopoulos
OMAC: A Holistic Optimization Framework for LLM-Based Multi-Agent Collaboration Shijun Li, Hilaf Hasson, Joydeep Ghosh
SoftJAX & SoftTorch: Empowering Automatic Differentiation Libraries with Informative Gradients Anselm Paulus, Andreas René Geist, Vit Musil et al.
← Browse the whole archive