Fixed-Point RNNs: Interpolating from Diagonal to Dense
NeurIPSSpotlight2025
TL;DR
Linear recurrent neural networks (RNNs) and state-space models (SSMs) such as Mamba have become promising alternatives to softmax-attention as sequence mixing layers in Transformer architectures…
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
transformer attention mamba
← All NeurIPS 2025 Spotlight papers · Browse the whole archive