Cheap to Hypothesize, Costly to Verify: The Defense Surface of Agentic Vulnerability Discovery
cs.CR, cs.AI
Submitted: 2026-09-28
Updated: 2026-10-07
Project page: https://xxbai.space/redherring
Terminology
Sources
- Trust Me, I Know This Function: Hijacking LLM Static Analysis using Bias
- DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence
- Clawdrain: Exploiting Tool-Calling Chains for Stealthy Token Exhaustion in OpenClaw Agents
- GLM-5: from Vibe Coding to Agentic Engineering
- Chaff Bugs: Deterring Attackers by Making Software Buggier
- Kimi K3: Open Frontier Intelligence
- CoTDeceptor:Adversarial Code Obfuscation Against CoT-Enhanced LLM Code Agents
- A Systematic Study of Code Obfuscation Against LLM-based Vulnerability Detection
- OTora: A Unified Red Teaming Framework for Reasoning-Level Denial-of-Service in LLM Agents
- Hacking Back the AI-Hacker: Prompt Injection as a Defense Against LLM-driven Cyberattacks
- LLM Agent Honeypot: Monitoring AI Hacking Agents in the Wild
- On Secure and Usable Program Obfuscation: A Survey
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs