Scal3R Reduces 3D Reconstruction Drift via Multi-Reference Pose Querying
September 2, 2026
Scal3R addresses geometric collapse in long-video 3D reconstruction by decoupling local geometry from global pose. It uses lightweight learnable tokens, representing 1% of parameters, to query poses relative to multiple past keyframes via asymmetric attention in a frozen backbone.
HOW THIS AFFECTS YOU
●
researcherYou can achieve more stable online reconstruction without retraining the entire backbone.