Formal Runtime Verification for Tool-Using LLM Agents: An Offline Same-Benchmark Study on AgentDojo and STAC
cs.CR
Submitted: 2026-10-07
Updated: 2026-10-07
Code: https://github.com/nikos-kekatos/formal-rv-tool-usingllm-agents
Terminology
Sources
- Maris: A Formally Verifiable Privacy Policy Enforcement Paradigm for Multi-Agent Collaboration Systems
- Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
- Mission-Level Runtime Assurance for LLM-Assisted ISR Swarms over a Verification-Aware Fabric
- SHADE-Arena: Evaluating Sabotage and Monitoring in LLM Agents
- STAC: When Innocent Tools Form Dangerous Chains for LLM Agents
- Progent: Securing AI Agents with Privilege Control
- GuardAgent: Safeguard LLM Agents by a Guard Agent via Knowledge-Enabled Reasoning
- Agent-SafetyBench: Evaluating the Safety of LLM Agents
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs