Diffusion Large Language Models for Visual Speech Recognition

arXiv:2605.28456 · cs.AI, cs.CV, eess.AS · Submitted 2026-05-27 · Read on arXiv

cs.AI, cs.CV, eess.AS

Submitted: 2026-05-27

Updated: 2026-09-01

Comments: Accepted to EMNLP 2026. Code: https://github.com/JeongHun0716/dllm-vsr

Code: https://github.com/JeongHun0716/dllm-vsr

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers