FOCAL-VLA: Subtask-Guided Geometry Distillation and Implicit World Modeling for Vision-Language-Action Models

arXiv:2609.21228 · cs.RO, cs.AI, cs.CV · Submitted 2026-09-18 · Read on arXiv

cs.RO, cs.AI, cs.CV

Submitted: 2026-09-18

Updated: 2026-09-18

Project page: https://zhiyuan-gao.github.io/FOCAL-VLA

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Related papers