TY - RPRT TI - On the Approximation and Convergence of Distributional Policy Gradient Algorithms for Risk-Sensitive Reinforcement Learning AU - Xian Yu AU - Minheng Xiao AU - Lei Ying PY - 2026 UR - https://arxiv.org/abs/2405.14749 ID - 2405.14749 ER -