MLLMCLIP: Feature-Level Distillation of MLLM for Robust Vision-Language Representations

arXiv:2608.25575 · cs.CV, cs.AI · Submitted 2026-08-26 · Read on arXiv

cs.CV, cs.AI

Submitted: 2026-08-26

Updated: 2026-08-26

Comments: EMNLP 2026 Main Conference

Code: https://github.com/LijieFan/LaCLIP

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers