Who Holds the Pen? Let Specifications, Not Agents, Sign Off
cs.AI, cs.MA
Submitted: 2026-09-24
Updated: 2026-09-24
Terminology
Sources
- Constitutional AI: Harmlessness from AI Feedback
- MCP-Atlas: A Large-Scale Benchmark for Tool-Use Competency with Real MCP Servers
- AgentRx: Diagnosing AI Agent Failures from Execution Trajectories
- AgentOps: Enabling Observability of LLM Agents
- Consensus on Transaction Commit
- SkillsBench: Benchmarking How Well Agent Skills Work Across Diverse Tasks
- VIGIL: Runtime Enforcement of Behavioral Specifications in AI Agent Skills
- Cryptographic Runtime Governance for Autonomous AI Systems: The Aegis Architecture for Verifiable Policy Enforcement
- Verify-Gated Completion as Admission Control in a Governed Multi-Agent Runtime: A Bounded Architecture Case Study
- Training language models to follow instructions with human feedback
- Formal Policy Enforcement for Real-World Agentic Systems
- Neuro-Symbolic Verification on Instruction Following of LLMs
- Verification-Gated Agentic Mission-State Governance for Intelligent Industrial Multi-Robot Systems
- Voyager: An Open-Ended Embodied Agent with Large Language Models
- AgentSpec: Customizable Runtime Enforcement for Safe and Reliable LLM Agents
- DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
- GLM-5: from Vibe Coding to Agentic Engineering
Related papers
- MAVEN-T: Reinforced Heterogeneous Distillation for Real-Time Multi-Agent Trajectory Prediction
- Model Discovery Agent: LLM-assisted Bayesian experiment design for data-efficient discovery of mechanistic world models
- The Clinician's Veto: Navigating Trust, Liability, and Uncertainty in Autonomous AI Prescribing
- MindHelper: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention
- Incumbent Advantage: Brand Bias and Cognitive Manipulation Dynamics in LLM Recommendation Systems
- VSAL: A Vision Solver with Adaptive Layouts for Graph Property Detection