NetInjectBench: Benchmarking Indirect Prompt Injection in Tool-Using Large Language Model Agents for Network Operations
cs.CR, cs.LG
Submitted: 2026-07-11
Updated: 2026-09-27
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Indirect Prompt Injections: Are Firewalls All You Need, or Stronger Benchmarks?
- Artificial Intelligence for IT Operations (AIOPS) Workshop White Paper
- On the Opportunities and Risks of Foundation Models
- StruQ: Defending Against Prompt Injection with Structured Queries
- The Llama 3 Herd of Models
- Defending Against Indirect Prompt Injection Attacks With Spotlighting
- Mistral 7B
- MRKL Systems: A modular, neuro-symbolic architecture that combines large language models, external knowledge sources and discrete reasoning
- AgentDyn: Are Your Agent Security Defenses Deployable in Real-World Dynamic Environments?
- Prompt Injection attack against LLM-integrated Applications
- A Systematic Mapping Study in AIOps
- Toolformer: Language Models Can Teach Themselves to Use Tools
- Model evaluation for extreme risks
- The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions
- ICON: Indirect Prompt Injection Defense for Agents based on Inference-Time Correction
- Ethical and social risks of harm from Language Models
- Qwen2.5 Technical Report
- Defense Against Indirect Prompt Injection via Tool Result Parsing
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs