Universal Pose Pretraining for Generalizable Vision-Language-Action Policies

arXiv:2602.19710 · cs.CV, cs.LG, cs.RO · Submitted 2026-02-23 · Read on arXiv

cs.CV, cs.LG, cs.RO

Submitted: 2026-02-23

Updated: 2026-09-27

Comments: Accepted to Robotics: Science and Systems (RSS) 2026. Project website: https://hetolin.github.io/PoseVLA

Journal ref: Robotics: Science and Systems, 2026

Code: https://github.com/QwenLM/Qwen3-VL

Project page: https://hetolin.github.io/PoseVLA

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers