Researcher profile

Xiaoyang Tan

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Smoothing Advantage Learning

    2022 · arXiv (Cornell University)

    Advantage learning (AL) aims to improve the robustness of value-based reinforcement learning against estimation errors with action-gap-based regularization. Unfortunately, the method tends to be unstable in the case of function approximation. In this paper, we …