Prompt Injection Detection for Email Agents Through Attack Chain Modeling
cs.CR, cs.CL
Submitted: 2026-09-25
Updated: 2026-09-25
Code: https://github.com/AcaiLab/Email
Terminology
Sources
- Tensor Trust: Interpretable Prompt Injection Attacks from an Online Game
- InjecGuard: Benchmarking and Mitigating Over-defense in Prompt Injection Guardrail Models
- Defeating Prompt Injections by Design
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs