Before Agents Speak: Pre-hoc Failure Risk Inference in Multi-Agent Systems
Shi Lin, Chenpei Wang, Peng Qian, Dezhang Kong, Minghao Li, Yufeng Li, Xun Wang
cs.CR
Submitted: 2026-07-29
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- GPT-4 Technical Report
- AutoHall: Automated Factuality Hallucination Dataset Generation for Large Language Models
- MedSentry: Understanding and Mitigating Safety Risks in Medical LLM Multi-Agent Systems
- The Llama 3 Herd of Models
- Towards a Science of Scaling Agent Systems
- A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
- Reasoning as a Weapon: Adaptive Dual-Path Jailbreak Attack on Large Language Models
- Qwen2 Technical Report
- A Stitch in Time Saves Nine: Detecting and Mitigating Hallucinations of LLMs by Validating Low-Confidence Generation
- Pushing the Limits of Safety: A Technical Report on the ATLAS Challenge 2025
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs