Emergent Risks in Generative Multi-Agent Systems
cs.MA, cs.CL, cs.CY
Submitted: 2026-03-29
Updated: 2026-09-10
License: http://creativecommons.org/licenses/by/4.0/
The gist: Multi-agent systems composed of large generative models are rapidly moving from laboratory prototypes to real-world deployments, where they jointly plan, negotiate, and allocate shared resources to
Terminology
Abstract
Multi-agent systems composed of large generative models are rapidly moving from laboratory prototypes to real-world deployments, where they jointly plan, negotiate, and allocate shared resources to solve complex tasks. While such systems promise unprecedented scalability and autonomy, their collective interaction also gives rise to failure modes that cannot be reduced to individual agents. Understanding these emergent risks is therefore critical. Here, we present a pioneer study of such emergent multi-agent risk in workflows that involve competition over shared resources (e.g., computing resources or market share), sequential handoff collaboration (where downstream agents see only predecessor outputs), collective decision aggregation, and others. Across these settings, we observe that such group behaviors arise frequently across repeated trials and a wide range of interaction conditions, rather than as rare or pathological cases. In particular, phenomena such as collusion-like coordination and conformity emerge with non-trivial frequency under realistic resource constraints, communication protocols, and role assignments, mirroring well-known pathologies in human societies despite no explicit instruction. Moreover, these risks cannot be prevented by existing agent-level safeguards alone. These findings expose the dark side of intelligent multi-agent systems: a social intelligence risk where agent collectives, despite no instruction to do so, spontaneously reproduce familiar failure patterns from human societies.
Sources
- Self-Resource Allocation in Multi-Agent LLM Systems
- Conformity and Social Impact on AI Agents
- Supracompetitive Pricing Under AI Monoculture
- Scheming AIs: Will AIs fake alignment during training in order to get power?
- Why Do Multi-Agent LLM Systems Fail?
- ChatEval: Towards Better LLM-based Evaluators through Multi-Agent Debate
- TravelAgent: An AI Assistant for Personalized Travel Planning
- AI4Research: A Survey of Artificial Intelligence for Scientific Research
- Artificial Intelligence and Algorithmic Price Collusion in Two-sided Markets
- Belief in Authority: Impact of Authority in Multi-Agent Evaluation Framework
- The Traitors: Deception and Trust in Multi-Agent Language Model Simulations
- A Review of Cooperation in Multi-agent Learning
- HonestLLM: Toward an Honest and Helpful Large Language Model
- Are Your Agents Upward Deceivers?
- Multi-Agent Risks from Advanced AI
- Position: Towards a Responsible LLM-empowered Multi-Agent Systems
- On the Resilience of LLM-Based Multi-Agent Collaboration with Faulty Agents
- On the Trustworthiness of Generative Foundation Models: Guideline, Assessment, and Perspective
- Deep Research Agents: A Systematic Examination And Roadmap
- Flooding Spread of Manipulated Knowledge in LLM-Based Multi-Agent Communities
Related papers
- Highway Congestion Reduction through Reinforcement Learning Based Eulerian Headway Control
- You Only Align Once: Propagating Cooperative Behaviors in Multi-Agent Systems through Seed Agents
- Deny Without Disabling: Authorization-Paired Evaluation and Control for Multi-Agent Systems
- MA-SAPO: Multi-Agent Reasoning for Score-Aware Prompt Optimization
- PeroMAS: A Multi-agent System of Perovskite Material Discovery
- StitchCUDA: An Automated Multi-Agents End-to-End GPU Programing Framework with Rubric-based Agentic Reinforcement Learning