WAInjectBench: Benchmarking Prompt Injection Detections for Web Agents
cs.CR, cs.AI, cs.CL
Submitted: 2025-10-01
Updated: 2026-09-23
Code: https://github.com/Norrrrrrr-lyn/WAInjectBench
Terminology
Sources
- Embedding-based classifiers can detect prompt injection attacks
- VPI-Bench: Visual Prompt Injection Attacks for Computer-Use Agents
- WASP: Benchmarking Web Agent Security Against Prompt Injection Attacks
- AugMix: A Simple Data Processing Method to Improve Robustness and Uncertainty
- Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
- VisualWebArena: Evaluating Multimodal Agents on Realistic Visual Web Tasks
- EIA: Environmental Injection Attack on Generalist Web Agents for Privacy Leakage
- Improving Robustness Without Sacrificing Accuracy with Patch Gaussian Augmentation
- PromptArmor: Simple yet Effective Prompt Injection Defenses
- Intriguing properties of neural networks
- MiniLMv2: Multi-Head Self-Attention Relation Distillation for Compressing Pretrained Transformers
- WebInject: Prompt Injection Attack to Web Agents
- Dissecting Adversarial Robustness of Multimodal LM Agents
- Feature Squeezing: Detecting Adversarial Examples in Deep Neural Networks
- JailGuard: A Universal Detection Framework for LLM Prompt-based Attacks
- Attacking Vision-Language Computer Agents via Pop-ups
- WebArena: A Realistic Web Environment for Building Autonomous Agents
- MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs