Share your thoughts, 1 month free Claude Pro on usSee more
WorkDL logo mark

DF-VO: What Should Be Learnt for Visual Odometry?

About

Multi-view geometry-based methods dominate the last few decades in monocular Visual Odometry for their superior performance, while they have been vulnerable to dynamic and low-texture scenes. More importantly, monocular methods suffer from scale-drift issue, i.e., errors accumulate over time. Recent studies show that deep neural networks can learn scene depths and relative camera in a self-supervised manner without acquiring ground truth labels. More surprisingly, they show that the well-trained networks enable scale-consistent predictions over long videos, while the accuracy is still inferior to traditional methods because of ignoring geometric information. Building on top of recent progress in computer vision, we design a simple yet robust VO system by integrating multi-view geometry and deep learning on Depth and optical Flow, namely DF-VO. In this work, a) we propose a method to carefully sample high-quality correspondences from deep flows and recover accurate camera poses with a geometric module; b) we address the scale-drift issue by aligning geometrically triangulated depths to the scale-consistent deep depths, where the dynamic scenes are taken into account. Comprehensive ablation studies show the effectiveness of the proposed method, and extensive evaluation results show the state-of-the-art performance of our system, e.g., Ours (1.652%) v.s. ORB-SLAM (3.247%}) in terms of translation error in KITTI Odometry benchmark. Source code is publicly available at: \href{https://github.com/Huangying-Zhan/DF-VO}{DF-VO}.

Huangying Zhan, Chamara Saroj Weerasekera, Jia-Wang Bian, Ravi Garg, Ian Reid• 2021

Related benchmarks

TaskDatasetResultRank
Visual OdometryKITTI Seq. 10
Translational Error (%)1.82
35
Visual OdometryOxford
Trajectory Error (11-12)2.03
18
Visual OdometryZJH-VO multi-scale (F-09)
ATE (m)10.56
18
Visual OdometryZJH-VO multi-scale (F-04)
ATE (m)11.96
18
Visual OdometryZJH-VO multi-scale (F-00)
ATE (m)18.94
18
Visual OdometryZJH-VO Average multi-scale
ATE (m)16.37
18
Visual OdometryZJH-VO multi-scale (F-01)
ATE (m)24.02
18
Visual OdometryNCLT
Sequence 03-17 Error25.52
16
Visual OdometryNCLT
FPS9.37
11
Showing 9 of 9 rows

Other info

Follow for update