Offline Guidance, Online Reasoning: Reusing LLM Feedback for Small Language Models

arXiv:2609.39346 · cs.CL · Submitted 2026-09-30 · Read on arXiv

cs.CL

Submitted: 2026-09-30

Updated: 2026-09-30

Code: https://github.com/ZBH031/reusable-latent-correction

Terminology

Sources

Related papers