ReFIne: A Framework for Trustworthy Large Reasoning Models with Reliability, Faithfulness, and Interpretability

arXiv:2510.09062 · cs.CL · Submitted 2025-10-10 · Read on arXiv

cs.CL

Submitted: 2025-10-10

Updated: 2026-08-25

Comments: Accpepted to COLM 2026

Code: https://github.com/Trustworthy-ML-Lab/Training_Trustworthy_LRM_with_Refine

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers