When Prompts Control Robots: Prompt Injection Attacks in Multi-Agent Robotic Systems
Neha Nagaraja, Amisha Bagari, Hayretdin Bahsi
cs.RO, cs.AI, cs.CR, cs.MA
Submitted: 2026-08-01
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Ignore Previous Prompt: Attack Techniques For Language Models
- Multi-Agent Collaboration Mechanisms: A Survey of LLMs
- A Survey of LLM-Driven AI Agent Communication: Protocols, Security Risks, and Defense Countermeasures
- Prompt Injection Attack to Tool Selection in LLM Agents
- Large Language Model based Multi-Agents: A Survey of Progress and Challenges
- IP Leakage Attacks Targeting LLM-Based Multi-Agent Systems
- Multi-Agent Systems Execute Arbitrary Malicious Code
- BadRobot: Jailbreaking Embodied LLM Agents in the Physical World
- Agent Smith: A Single Image Can Jailbreak One Million Multimodal LLM Agents Exponentially Fast
- Automatic and Universal Prompt Injection Attacks against Large Language Models
Related papers
- FMT x: An Efficient and Asymptotically Optimal Extension of the Fast Marching Tree for Dynamic Replanning
- MPCFormer: A physics-informed data-driven approach for explainable socially-aware autonomous driving
- RoboLab: A High-Fidelity Simulation Benchmark for Analysis of Task Generalist Policies
- HRDexDB: A 4D Dexterous Grasping Dataset Across Human and Multiple Robot Embodiments
- APT: Action Expert Pretraining Improves Instruction Generalization of Vision-Language-Action Policies
- Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving