ExStereo: Lifting 2D Vision-Language-Action Models to 3D with Explicit Stereo Representations

arXiv:2610.04805 · cs.RO, cs.CV · Submitted 2026-10-03 · Read on arXiv

cs.RO, cs.CV

Submitted: 2026-10-03

Updated: 2026-10-03

Related papers