OverAct: Measuring and Mitigating Proactive Over-Authorization in LLM Tool-Calling Agents
cs.CR, cs.CL
Submitted: 2026-10-01
Updated: 2026-10-01
Terminology
Sources
- Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
- DeepSeek-V3 Technical Report
- DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
- Sycophancy to Subterfuge: Investigating Reward-Tampering in Large Language Models
- ToolTalk: Evaluating Tool-Usage in a Conversational Setting
- Prompt Flow Integrity to Prevent Privilege Escalation in LLM Agents
- A Vision for Access Control in LLM-based Agent Systems
- Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks
- How RLHF Amplifies Sycophancy
- Progent: Securing AI Agents with Privilege Control
- Kimi K2: Open Agentic Intelligence
- Qwen3 Technical Report
- ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
- The Tool-Overuse Illusion: Why Does LLM Prefer External Tools over Internal Knowledge?
- Agent-SafetyBench: Evaluating the Safety of LLM Agents
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs