ReBeCA: Unveiling Interpretable Behavior Hierarchy behind the Iterative Self-Reflection of Language Models with Causal Analysis

arXiv:2602.06373 · cs.CL · Submitted 2026-02-06 · Read on arXiv

cs.CL

Submitted: 2026-02-06

Updated: 2026-09-06

Comments: Accepted to Findings of the Association for Computational Linguistics: EMNLP 2026

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers