Human Label Variation as Stable Signal: Learning Annotator-Specific Explanation Behavior via Cross-Annotator Preference Optimization
cs.CL
Submitted: 2026-05-27
Updated: 2026-08-28
Code: https://github.com/mainlp/CAPO
Terminology
Sources
- Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective
- Out of One, Many: Using Language Models to Simulate Human Samples
- ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks
- Aligning LLM Uncertainty with Human Disagreement in Subjectivity Analysis
- Agree, Disagree, Explain: Decomposing Human Label Variation in NLI through the Lens of Explanations
- Entropy and type-token ratio in gigaword corpora
- The Llama 3 Herd of Models
- Qwen3 Technical Report
- Steering Language Models With Activation Engineering
- EVADE: LLM-Based Explanation Generation and Validation for Error Detection in NLI
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering