PhysMLLMs: Spatial Priors for Unified Referring Segmentation and Grounded Reasoning of Images and Videos

arXiv:2608.24574 · cs.AI · Submitted 2026-08-25 · Read on arXiv

cs.AI

Submitted: 2026-08-25

Updated: 2026-08-25

Code: https://github.com/tusu-code/20260121-icml2026-2

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers