Value Mirror Descent for Reinforcement Learning

arXiv:2604.06039 · math.OC, cs.LG, math.PR · Submitted 2026-04-07 · Read on arXiv

math.OC, cs.LG, math.PR

Submitted: 2026-04-07

Updated: 2026-09-23

Terminology

Sources

Related papers