CodePoisonRAG: Knowledge Poisoning Attacks on Retrieval-Augmented Code Generation
cs.CR, cs.LG
Submitted: 2026-09-02
Updated: 2026-09-02
Project page: https://cwe.mitre.org/data/definitions/79.html
Terminology
Sources
- RAVEN: Agentic RAG for Automated Vulnerability Repair
- Security and Quality in LLM-Generated Code: A Multi-Language, Multi-Model Analysis
- RESCUE: Retrieval Augmented Secure Code Generation
- Code Llama: Open Foundation Models for Code
- DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
- GLM-5: from Vibe Coding to Agentic Engineering
- jina-embeddings-v3: Multilingual Embeddings With Task LoRA
- The Faiss library
- CodeBLEU: a Method for Automatic Evaluation of Code Synthesis
- CYBERSECEVAL 3: Advancing the Evaluation of Cybersecurity Risks and Capabilities in Large Language Models
- Retrieval-Augmented Generation for Large Language Models: A Survey
- Backdoored Retrievers for Prompt Injection Attacks on Retrieval Augmented Generation of Large Language Models
- MM-PoisonRAG: Disrupting Multimodal RAG with Local and Global Poisoning Attacks
- Retrieval-Augmented Code Generation: A Survey with Focus on Repository-Level Approaches
- RAG-Pull: Turning Retrieval into a Code-Injection Channel via Invisible Unicode Perturbations
- SOSecure: Safer Code Generation with RAG and StackOverflow Discussions
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs