Learning Evidence Highlighting for Frozen LLMs
cs.CL, cs.AI
Submitted: 2026-04-24
Updated: 2026-09-26
Terminology
Sources
- PRL: Prompts from Reinforcement Learning
- OneRec: Unifying Retrieve and Rank with Generative Recommender and Iterative Preference Alignment
- Chat-REC: Towards Interactive and Explainable LLMs-Augmented Recommender System
- Gemma 3 Technical Report
- The Llama 3 Herd of Models
- Bridging Language and Items for Retrieval and Recommendation: Benchmarking LLMs as Semantic Encoders
- GR2: Generative Reasoning Re-ranker
- OneRec-Think: In-Text Reasoning for Generative Recommendation
- System 2 Attention (is something you might need too)
- On the Generalization of SFT: A Reinforcement Learning Perspective with Reward Rectification
- Qwen3 Technical Report
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering