Enhancing Assessment of Self-Consistency in LLM Explanations using Perturbation Strength
cs.CL
Submitted: 2026-09-25
Updated: 2026-09-25
Terminology
Sources
- Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
- A Survey on Data Contamination for Large Language Models
- Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
- The Llama 3 Herd of Models
- Mistral 7B
- Measuring Faithfulness in Chain-of-Thought Reasoning
- Noiser: Bounded Input Perturbations for Attributing Large Language Models
- MiniLM: Deep Self-Attention Distillation for Task-Agnostic Compression of Pre-Trained Transformers
- On the Predictive Power of Neural Language Models for Human Real-Time Comprehension Behavior
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering