Calibration, Not Answer Selection: Distilling Internal Confidence in Reasoning Models

arXiv:2609.33290 · cs.CL · Submitted 2026-09-27 · Read on arXiv

cs.CL

Submitted: 2026-09-27

Updated: 2026-09-27

Code: https://github.com/yale-nlp/RLMF

Terminology

Sources

Related papers