AgentSnare: Learning to Delay, Divert, and Defuse Autonomous Penetration Agents
cs.CR, cs.CL, cs.LG
Submitted: 2026-07-29
Updated: 2026-09-30
Code: https://github.com/cowrie/cowrie
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- FireAct: Toward Language Agent Fine-tuning
- VulnBot: Autonomous Penetration Testing for A Multi-Agent Collaborative Framework
- Can We Stop Malicious AI? KILLBENCH: A Benchmark for External AI Kill Switch Feasibility
- Hacking Back the AI-Hacker: Prompt Injection as a Defense Against LLM-driven Cyberattacks
- DARWIN: Evolving Jailbreak Adversary and Guardrail for LLM Safety Evaluation and Protection
- Dynamic Jailbreaking Attack
- Qwen3 Technical Report
- From Topology to Behavioral Semantics: Enhancing BGP Security by Understanding BGP's Language with LLMs
- CVE-Bench: A Benchmark for AI Agents' Ability to Exploit Real-World Web Application Vulnerabilities
- Cyber-Zero: Training Cybersecurity Agents without Runtime
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs