VR-JEPA: Learning Contrastive-State Latent Guidance for Generation-based Video Reasoning

arXiv:2609.40129 · cs.CV · Submitted 2026-09-30 · Read on arXiv

cs.CV

Submitted: 2026-09-30

Updated: 2026-09-30

Code: https://github.com/Wan-Video/Wan2.2

Terminology

Sources

Related papers