Scaling Model-Generated Distillation Data Can Make Latent Teacher Traits More Recoverable

arXiv:2608.26958 · cs.LG, cs.CL · Submitted 2026-08-27 · Read on arXiv

cs.LG, cs.CL

Submitted: 2026-08-27

Updated: 2026-08-27

Code: https://github.com/openai/grade-school-mathhttps:

Terminology

Sources

Related papers