arxivstat.MLcs.LG2026-07-09
Statistical Efficiency and Inference of Quantile Distributional Reinforcement Learning
Zijie Cheng, Yang Peng, Zhihua Zhang
In this paper, we study quantile-based distributional reinforcement learning from the perspective of statistical efficiency. We focus on distributional policy evaluation, whose goal is to characterize the return distribution, namely the distribution of discounted cumulative rewar…