CORTEXA
← Browse

Yang Peng

1 paper indexed

arxivstat.MLcs.LG2026-07-09

Statistical Efficiency and Inference of Quantile Distributional Reinforcement Learning

Zijie Cheng, Yang Peng, Zhihua Zhang

In this paper, we study quantile-based distributional reinforcement learning from the perspective of statistical efficiency. We focus on distributional policy evaluation, whose goal is to characterize the return distribution, namely the distribution of discounted cumulative rewar…

View free PDFSource page