Making LLMs Say What They Think: Measuring and Improving CoT-Interpretability Alignment

arXiv:2609.38972 · cs.CL, cs.AI · Submitted 2026-09-30 · Read on arXiv

cs.CL, cs.AI

Submitted: 2026-09-30

Updated: 2026-09-30

Code: https://github.com/yihuaihong/CIA-minimal-repro

Terminology

Sources

Related papers