The Detectability Gap: Hidden Heterogeneity in Hallucination Detection Across Language Models
cs.CL, cs.AI
Submitted: 2026-09-25
Updated: 2026-09-25
Terminology
Sources
- TraceDet: Hallucination Detection from the Decoding Trace of Diffusion Large Language Models
- The Llama 3 Herd of Models
- TDGNet: Hallucination Detection in Diffusion Language Models via Temporal Dynamic Graphs
- SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models
- Large Language Diffusion Models
- DynHD: Hallucination Detection for Diffusion Large Language Models via Denoising Dynamics Deviation Learning
- Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks
- Qwen2.5 Technical Report
- Self-Consistency Improves Chain of Thought Reasoning in Language Models
- Dream 7B: Diffusion Large Language Models
- SAC3: Reliable Hallucination Detection in Black-Box Language Models via Semantic-aware Cross-check Consistency
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering