The Alignment Paradox: How Post-Training Amplifies Confident Hallucinations in Language Models

arXiv:2609.32617 · cs.CL, cs.AI · Submitted 2026-09-26 · Read on arXiv

cs.CL, cs.AI

Submitted: 2026-09-26

Updated: 2026-09-26

Code: https://github.com/star5o/HCE

Terminology

Sources

Related papers