Train Overcomplete, Deploy Compact: Scaling Recovery Capacity for Structured LLM Pruning

arXiv:2609.06974 · cs.CL, cs.AI · Submitted 2026-09-07 · Read on arXiv

cs.CL, cs.AI

Submitted: 2026-09-07

Updated: 2026-09-07

Comments: Accepted by the Conference on Empirical Methods in Natural Language Processing (EMNLP) 2026

Code: https://github.com/mmai-laboratory/OverRep

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Related papers