CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models

arXiv:2602.17684 · cs.LG, cs.AI · Submitted 2026-02-04 · Read on arXiv

cs.LG, cs.AI

Submitted: 2026-02-04

Updated: 2026-09-27

Code: https://github.com/volcengine/verl

Project page: https://lark-ai-lab.github.io/codescaler.github.io

Terminology

Sources

Related papers