VidMap: Exploiting Temporal Structure for Video-Based Structure-from-Motion
- Published
- Source
- arXiv
- Paper number
- 764
- Field
- Computer Vision
- arXiv ID
- 2607.27194
Key points
- We combine SLAM's sequential tracking with SfM's global optimization to overcome the limitations of both.
- We use temporal order as the primary signal to prevent visual aliasing errors.
- We add a monocular depth prior to global optimization to address scale drift and degenerate configurations.
- We outperform existing SLAM, SfM, and learning-based methods by a wide margin on a challenging dataset with extreme motion.
Paper links
External research summaries. These are not HDATF publications or measured product results.