Researcher profile
Xiaoyang Tan
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Smoothing Advantage Learning
2022 · arXiv (Cornell University)
Advantage learning (AL) aims to improve the robustness of value-based reinforcement learning against estimation errors with action-gap-based regularization. Unfortunately, the method tends to be unstable in the case of function approximation. In this paper, we …