MiMo-V2.6: Scaling Reinforcement Learning Towards Self-Improvement

arXiv:2610.11959 · cs.CL · Submitted 2026-10-08 · Read on arXiv

cs.CL

Submitted: 2026-10-08

Updated: 2026-10-08

Code: https://github.com/datacurve-ai/deep-swe

Terminology

Sources

Related papers