VidMap: Exploiting Temporal Structure for Video-Based Structure-from-Motion

Published
Source
arXiv
Paper number
764
Field
Computer Vision
arXiv ID
2607.27194

Key points

  • We combine SLAM's sequential tracking with SfM's global optimization to overcome the limitations of both.
  • We use temporal order as the primary signal to prevent visual aliasing errors.
  • We add a monocular depth prior to global optimization to address scale drift and degenerate configurations.
  • We outperform existing SLAM, SfM, and learning-based methods by a wide margin on a challenging dataset with extreme motion.

Paper links

External research summaries. These are not HDATF publications or measured product results.

Read original (opens in a new tab)