Train at Moving Edge: Online-Verified Prompt Selection for Efficient RL Training of Large Reasoning Model

arXiv:2603.25184 · cs.LG, cs.AI · Submitted 2026-03-26 · Read on arXiv

cs.LG, cs.AI

Submitted: 2026-03-26

Updated: 2026-09-02

Code: https://github.com/huggingface/open-r1

Terminology

Sources

Related papers