Lightning OPD: Efficient Post-Training for Large Reasoning Models with Offline On-Policy Distillation

arXiv:2604.13010 · cs.LG, cs.AI · Submitted 2026-04-14 · Read on arXiv

cs.LG, cs.AI

Submitted: 2026-04-14

Updated: 2026-09-26

Code: https://github.com/jet-ai-projects/Lightning-OPD

Terminology

Sources

Related papers