Beyond Refusal Patterns: Safe-Role Internalization for Robust and Generalizable LLM Safety Alignment

arXiv:2610.07023 · cs.AI, cs.CL, cs.IR · Submitted 2026-10-04 · Read on arXiv

cs.AI, cs.CL, cs.IR

Submitted: 2026-10-04

Updated: 2026-10-04

Terminology

Related papers