CritiQ: Mining Data Quality Criteria from Human Preferences
cs.CL
Submitted: 2025-02-26
Updated: 2025-09-11
Comments: to be published in ACL 2025, Code is available at https://github.com/KYLN24/CritiQ
Journal ref: https://aclanthology.org/2025.acl-long.792/
DOI: 10.18653/v1/2025.acl-long.792
Code: https://github.com/KYLN24/CritiQ
License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/
Terminology
Sources
- Program Synthesis with Large Language Models
- InternLM2 Technical Report
- Self-RAG: Learning to Retrieve, Generate, and Critique through Self-Reflection
- Evaluating Large Language Models Trained on Code
- DeepSeek-V3 Technical Report
- Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
- Training Verifiers to Solve Math Word Problems
- The Llama 3 Herd of Models
- Textbooks Are All You Need
- Measuring Mathematical Problem Solving With the MATH Dataset
- LongWanjuan: Towards Systematic Measurement for Long Text Quality
- OpenCoder: The Open Cookbook for Top-Tier Code Large Language Models
- Self-Refine: Iterative Refinement with Self-Feedback
- Large Language Models are Zero-Shot Reasoners
- When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale
- Stochastic Beams and Where to Find Them: The Gumbel-Top-k Trick for Sampling Sequences Without Replacement
- Pretraining Language Models with Human Preferences
- Decoupled Weight Decay Regularization
- StarCoder 2 and The Stack v2: The Next Generation
- GPT-4 Technical Report
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering