MVDG: Efficient Multi-view 3D Disambiguation on Unconstrained Real-World Images
cs.CV
Submitted: 2026-10-01
Updated: 2026-10-01
Terminology
Sources
- Doppelgangers: Learning to Disambiguate Images of Similar Structures
- SuperPoint: Self-Supervised Interest Point Detection and Description
- Grounding Image Matching in 3D with MASt3R
- SuperGlue: Learning Feature Matching with Graph Neural Networks
- FastVGGT: Training-Free Acceleration of Visual Geometry Transformer
- LoFTR: Detector-Free Local Feature Matching with Transformers
- 3D Reconstruction with Spatial Memory
- DUSt3R: Geometric 3D Vision Made Easy
- Doppelgangers++: Improved Visual Disambiguation with Geometric 3D Features
- Learning to Find Good Correspondences
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models