arXiv · 2010.04805
Discussion of Kallus (2020) and Mo, Qi, and Liu (2020): New Objectives for Policy Learning
Abstract
We discuss the thought-provoking new objective functions for policy learning that were proposed in "More efficient policy learning via optimal retargeting" by Nathan Kallus and "Learning optimal distributionally robust individualized treatment rules" by Weibin Mo, Zhengling Qi, and Yufeng Liu. We show that it is important to take the curvature of the value function into account when working within the retargeting framework, and we introduce two ways to do so. We also describe more efficient approaches for leveraging calibration data when learning distributionally robust policies.
Explore related subjects
Keep this discovery
Sijia Li, Xiudi Li, Alex Luedtke. 2020-10-09. Discussion of Kallus (2020) and Mo, Qi, and Liu (2020): New Objectives for Policy Learning. https://arxiv.org/abs/2010.04805
Cite the original work for its findings. Save a collection to share your selection of sources.