Auditable Agents
cs.AI
Submitted: 2026-04-07
Updated: 2026-09-27
Comments: 24 pages, 4 figures. Extended version. A condensed 4-page version appears in the Proceedings of the ACM AI Leadership Summit 2026 (Visionary Papers track)
Code: https://github.com/openclaw/openclaw
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- AgentRx: Diagnosing AI Agent Failures from Execution Trajectories
- AgentOps: Enabling Observability of LLM Agents
- MLAgentBench: Evaluating Language Agents on Machine Learning Experimentation
- Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
- Agent Safety Is Action Alignment
- The Autonomy Tax: Defense Training Breaks LLM Agents
- FORTIS: Benchmarking Over-Privilege in Agent Skills
- When Only the Final Text Survives: Implicit Execution Tracing for Multi-Agent Auditing
- Audit Trails for Accountability in Large Language Models
- Generative Agents: Interactive Simulacra of Human Behavior
- Gorilla: Large Language Model Connected with Massive APIs
- Fine-tuning Aligned Language Models Compromises Safety, Even When Users Do Not Intend To!
- ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs
- SkillCenter: A Large-Scale Source-Grounded Skill Library for Autonomous AI Agents
- Authenticated Delegation and Authorized AI Agents
- AgentSpec: Customizable Runtime Enforcement for Safe and Reliable LLM Agents
- OpenHands: An Open Platform for AI Software Developers as Generalist Agents
- Agent Skills for Large Language Models: Architecture, Acquisition, Security, and the Path Forward
- SWE-agent: Agent-Computer Interfaces Enable Automated Software Engineering
- No Attacker Needed: Unintentional Cross-User Contamination in Shared-State LLM Agents
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection