Speculative Safety Honeypot: Toward Proactive Defense Against Multi-turn Agent Attacks
cs.CR, cs.AI, cs.CL
Submitted: 2026-09-30
Updated: 2026-09-30
Terminology
Sources
- Dynamic Speculative Agent Planning
- Defending Against Indirect Prompt Injection Attacks With Spotlighting
- STAC: When Innocent Tools Form Dangerous Chains for LLM Agents
- Prompt Injection Attacks on Agentic Coding Assistants: A Systematic Analysis of Vulnerabilities in Skills, Tools, and Protocol Ecosystems
- SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning
- LLMs know their vulnerabilities: Uncover Safety Gaps through Natural Distribution Shifts
- PromptArmor: Simple yet Effective Prompt Injection Defenses
- MultiLoRA: Democratizing LoRA for Better Multi-Task Learning
- TOUCAN: Synthesizing 1.5M Tool-Agentic Data from Real-World MCP Environments
- Qwen3 Technical Report
- Defense Against Indirect Prompt Injection via Tool Result Parsing
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs