Sieve and Sage: Efficient Distraction Filtering for Reliable RALM Abstention
cs.CL, cs.LG
Submitted: 2026-09-17
Updated: 2026-09-17
Terminology
Sources
- Longformer: The Long-Document Transformer
- Question Answering with Subgraph Embeddings
- Med42 -- Evaluating Fine-Tuning Strategies for Medical LLMs: Full-Parameter vs. Parameter-Efficient Approaches
- Enhancing Noise Robustness of Retrieval-Augmented Language Models with Adaptive Adversarial Training
- The Llama 3 Herd of Models
- MASH: Modeling Abstention via Selective Help-Seeking
- LoRA: Low-Rank Adaptation of Large Language Models
- Sufficient Context: A New Lens on Retrieval Augmented Generation Systems
- RE-RAG: Improving Open-Domain QA Performance and Interpretability with Relevance Estimator in Retrieval-Augmented Generation
- Overview of BioASQ 2025: The Thirteenth BioASQ Challenge on Large-Scale Biomedical Semantic Indexing and Question Answering
- MedMCQA : A Large-scale Multi-Subject Multi-Choice Dataset for Medical domain Question Answering
- In-Context Retrieval-Augmented Language Models
- ColBERTv2: Effective and Efficient Retrieval via Lightweight Late Interaction
- DomainRAG: A Chinese Benchmark for Evaluating Domain-specific Retrieval-Augmented Generation
- Certifiably Robust RAG against Retrieval Corruption
- Corrective Retrieval Augmented Generation
- HotpotQA: A Dataset for Diverse, Explainable Multi-hop Question Answering
- Making Retrieval-Augmented Language Models Robust to Irrelevant Context
- Chain-of-Note: Enhancing Robustness in Retrieval-Augmented Language Models
- RAFT: Adapting Language Model to Domain Specific RAG
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering