Package Hallucination Attacks on Coding Agents through Prompt Injection in Rule Files
cs.CR, cs.AI
Submitted: 2026-10-07
Updated: 2026-10-07
Code: https://github.com/agentsmd/agents.md
Terminology
Sources
- Universal and Transferable Adversarial Attacks on Aligned Language Models
- OpenAI GPT-5 System Card
- Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
- Large Language Models Hallucination: A Comprehensive Survey
- Why Language Models Hallucinate
- Mechanisms of Prompt-Induced Hallucination in Vision-Language Models
- Qwen2.5-Coder Technical Report
- Qwen3 Technical Report
- Prompt Repetition Improves Non-Reasoning LLMs
- PromptArmor: Simple yet Effective Prompt Injection Defenses
- PIShield: Detecting Prompt Injection Attacks via Intrinsic LLM Features
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs