SpanCalib-VLM: Calibrated Hallucination Span Detection in Vision-Language Models
cs.CV, cs.CL
Submitted: 2026-08-30
Updated: 2026-08-30
Code: https://github.com/Aman-byte1/Hallucination-Detection-in-LVLMs
Terminology
Sources
- Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
- Thinking Fast and Slow in AI
- MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
- On Calibration of Modern Neural Networks
- Language Models (Mostly) Know What They Know
- From System 1 to System 2: A Survey of Reasoning Large Language Models
- A Survey on Hallucination in Large Vision-Language Models
- Can Humans Dream of Electric Sheep? Human-Written Samples for Fine-Grained Vision-and-Language Hallucination Benchmarking
- Can You Trust Your Model's Uncertainty? Evaluating Predictive Uncertainty Under Dataset Shift
- Qwen3 Technical Report
- MiniCPM-V: A GPT-4V Level MLLM on Your Phone
- MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe
Related papers
- Loss Knows Best: Detecting Annotation Errors in Videos via Loss Trajectories
- AnchorWeave: World-Consistent Video Generation with Retrieved Local Spatial Memories
- Benchmarking the Robustness of Foundation Models for Mammography under Domain Shift
- MambaX-Net: Dual-Input Mamba-Enhanced Cross-Attention Network for Longitudinal MRI Segmentation
- TeleOCR: Navigating Document Parsing Across Digital and Camera-Captured Documents
- A Survey on Efficient Vision-Language-Action Models