StarOR: Synergizing Tree Search and Test-Time Reinforcement Learning for Optimization Modeling

arXiv:2606.15197 · cs.LG, cs.AI · Submitted 2026-06-13 · Read on arXiv

cs.LG, cs.AI

Submitted: 2026-06-13

Updated: 2026-09-26

Code: https://github.com/volcengine/verl

Terminology

Sources

Related papers