Online3R: Online Learning for Consistent Sequential Reconstruction Based on Geometry Foundation Model

About

We present Online3R, a new sequential reconstruction framework that is capable of adapting to new scenes through online learning, effectively resolving inconsistency issues. Specifically, we introduce a set of learnable lightweight visual prompts into a pretrained, frozen geometry foundation model to capture the knowledge of new environments while preserving the fundamental capability of the foundation model for geometry prediction. To solve the problems of missing groundtruth and the requirement of high efficiency when updating these visual prompts at test time, we introduce a local-global self-supervised learning strategy by enforcing the local and global consistency constraints on predictions. The local consistency constraints are conducted on intermediate and previously local fused results, enabling the model to be trained with high-quality pseudo groundtruth signals; the global consistency constraints are operated on sparse keyframes spanning long distances rather than per frame, allowing the model to learn from a consistent prediction over a long trajectory in an efficient way. Our experiments demonstrate that Online3R outperforms previous state-of-the-art methods on various benchmarks. Project page: https://shunkaizhou.github.io/online3r-1.0/

Shunkai Zhou, Zike Yan, Fei Xue, Dong Wu, Yuchen Deng, Hongbin Zha• 2026

Related benchmarks

Task	Dataset	Result
3D Reconstruction	7 Scenes	Completion6.7	161
3D Reconstruction	NRGBD	--	66
Camera pose estimation	TUM RGB-D	Error (desk)0.016	15
Camera pose estimation	NRGBD	Error (Breakroom)0.102	5

Showing 4 of 4 rows

Other info

Follow for update

@wizwand_team Discord