Depth Anything 3: Recovering the Visual Space from Any Views
ICLROral2026
TL;DR
Depth Anything 3 uses a single vanilla DINOv2 transformer to take arbitrary input views and outputs consistent depth and ray maps, delivering leading pose, geometry, and visual rendering performance.
Opening excerpt from the authors’ abstract. source
Read the paper
Topics
transformer