Fixed-Point RNNs: Interpolating from Diagonal to Dense

NeurIPSSpotlight2025

Authors
Sajad Movahedi, Felix Sarnthein, Nicola Muca Cirone, Antonio Orvieto
Venue
NeurIPS 2025
Track
Spotlight

TL;DR

Linear recurrent neural networks (RNNs) and state-space models (SSMs) such as Mamba have become promising alternatives to softmax-attention as sequence mixing layers in Transformer architectures…

Opening excerpt from the authors’ abstract. source

Read the paper

Topics

transformer attention mamba

← All NeurIPS 2025 Spotlight papers · Browse the whole archive