Return or Revise? Learning When Revision Helps Retrieval-Augmented QA
cs.CL, cs.IR, cs.LG
Submitted: 2026-09-24
Updated: 2026-09-24
Code: https://github.com/openai/gpt-oss
Terminology
Sources
- MS MARCO: A Human Generated MAchine Reading COmprehension Dataset
- The Llama 3 Herd of Models
- Is Escalation Worth It? On the Depth of LLM Cascades
- CounterRefine: Answer-Conditioned Counterevidence Retrieval for Inference-Time Knowledge Repair in Factual Question Answering
- Language Models (Mostly) Know What They Know
- RASER: Recoverability-Aware Selective Escalation Router for Multi-Hop Question Answering
- Self-Correction as Feedback Control: Error Dynamics, Stability Thresholds, and Prompt Interventions in LLMs
- Signed Rescue Routing: Harm-Aware Cascades for Efficient LLM Inference
- gpt-oss-120b & gpt-oss-20b Model Card
- Olmo 3
- Trust or Abstain? A Self-Aware RAG Approach
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering