RL Starts before RL: On Policy Distillation for Better Reinforcement Learning

arXiv:2609.28145 · cs.LG · Submitted 2026-09-23 · Read on arXiv

cs.LG

Submitted: 2026-09-23

Updated: 2026-09-23

Terminology

Sources

Related papers