Frequency Matters: Fast Model-Agnostic Data Curation for Pruning and Quantization
cs.CL, cs.AI
Submitted: 2026-03-17
Updated: 2026-08-27
Code: https://github.com/FrancescoMonaco/ZipCal
Terminology
Sources
- Is C4 Dataset Optimal for Pruning? An Investigation of Calibration Data for LLM Pruning
- Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge
- Training Verifiers to Solve Math Word Problems
- An Efficient Matrix Multiplication Algorithm for Accelerating Inference in Binary and Ternary Neural Networks
- Measuring Mathematical Problem Solving With the MATH Dataset
- GPTQ: Accurate Post-Training Quantization for Generative Pre-trained Transformers
- The Pile: An 800GB Dataset of Diverse Text for Language Modeling
- The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits
- When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale
- SlimLLM: Accurate Structured Pruning for Large Language Models
- Pointer Sentinel Mixture Models
- Preserving LLM Capabilities through Calibration Data Curation: From Analysis to Optimization
- A Simple and Effective Pruning Approach for Large Language Models
- LUT-GEMM: Quantized Matrix Multiplication based on LUTs for Efficient Inference in Large-Scale Generative Language Models
- Are NLP Models really able to Solve Simple Math Word Problems?
- Gemma 2: Improving Open Language Models at a Practical Size
- BitNet: Scaling 1-bit Transformers for Large Language Models
- HellaSwag: Can a Machine Really Finish Your Sentence?
Related papers
- Exploring Solution Divergence and Its Effect on Large Language Model Problem Solving
- Ishigaki-IDS-Bench: A Benchmark for Generating Information Delivery Specification from BIM Information Requirements
- Subliminal Steering: Stronger Encoding of Hidden Signals
- MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
- The End of Transformers? On Challenging Attention and the Rise of Sub-Quadratic Architectures
- Untangling the Mechanisms of Misleading Context in Medical Question Answering