RAGSentinel: Certifiable Geometric Consensus for Robust Retrieval-Augmented Generation
cs.CR, cs.AI, cs.IR, cs.LG
Submitted: 2026-08-25
Updated: 2026-08-25
Comments: To appear in EMNLP 2026 (Main Conference)
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- GPT-4 Technical Report
- Is My Data in Your Retrieval Database? Membership Inference Attacks Against Retrieval Augmented Generation
- MS MARCO: A Human Generated MAchine Reading COmprehension Dataset
- Discovering Latent Knowledge in Language Models Without Supervision
- M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
- The Llama 3 Herd of Models
- Unsupervised Dense Information Retrieval with Contrastive Learning
- Mistral 7B
- Ignore Previous Prompt: Attack Techniques For Language Models
- Improving Text Embeddings with Large Language Models
- Certifiably Robust RAG against Retrieval Corruption
- Approximate Nearest Neighbor Negative Contrastive Learning for Dense Text Retrieval
- Benchmarking Poisoning Attacks against Retrieval-Augmented Generation
- TrustRAG: Enhancing Robustness and Trustworthiness in Retrieval-Augmented Generation
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs