Nonparametric Variance-Penalized Actor-Critic: Statistical Inference for Risk-Sensitive Reinforcement Learning

arXiv:2609.14327 · cs.LG, stat.ML · Submitted 2026-09-13 · Read on arXiv

cs.LG, stat.ML

Submitted: 2026-09-13

Updated: 2026-09-13

Comments: Submitted to IEEE Transactions on Neural Networks and Learning Systems. This work has been submitted to the IEEE for possible publication. Copyright may be transferred without notice, after which this version may no longer be accessible

License: http://arxiv.org/licenses/nonexclusive-distrib/1.0/

Terminology

Related papers