Knowledge-Verified Emergent Deception in LLM Agents Under Conflicting Incentives

arXiv:2608.26372 · cs.CL, cs.AI · Submitted 2026-08-26 · Read on arXiv

cs.CL, cs.AI

Submitted: 2026-08-26

Updated: 2026-08-26

Code: https://github.com/meta-llama/llama-models

Terminology

Sources

Related papers