Does Deeper Reasoning Compromise Alignment? Revealing and Mitigating of Alignment Collapse in Large Reasoning Models

arXiv:2609.08186 · cs.AI, cs.CL · Submitted 2026-09-08 · Read on arXiv

cs.AI, cs.CL

Submitted: 2026-09-08

Updated: 2026-09-08

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Related papers