Safe Score Matching: Diffusion Policies with Hamilton-Jacobi Reachability for Online Safe Reinforcement Learning

arXiv:2609.33337 · cs.LG, cs.RO · Submitted 2026-09-27 · Read on arXiv

cs.LG, cs.RO

Submitted: 2026-09-27

Updated: 2026-09-27

Code: https://github.com/byli888/safe-score-matching

Terminology

Sources

Related papers