Squeeze10-LLM: Squeezing LLMs' Weights by 10 Times via a Staged Mixed-Precision Quantization Method

arXiv:2507.18073 · cs.LG · Submitted 2025-07-24 · Read on arXiv

cs.LG

Submitted: 2025-07-24

Updated: 2026-09-08

Code: https://github.com/EleutherAI/lmevaluation-harness

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers