Train Where the Quantized Model Goes: On-Policy Distillation for Low-Bit Reasoning

arXiv:2609.26708 · cs.LG, cs.AI · Submitted 2026-09-22 · Read on arXiv

cs.LG, cs.AI

Submitted: 2026-09-22

Updated: 2026-09-22

Comments: 18 pages, 6 figures

Code: https://github.com/EleutherAI/lm-evaluation-harness

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Sources

Related papers