Researcher profile
Qingkai Liang
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Accelerated Primal-Dual Policy Optimization for Safe Reinforcement Learning
2018 · arXiv (Cornell University)
Constrained Markov Decision Process (CMDP) is a natural framework for reinforcement learning tasks with safety constraints, where agents learn a policy that maximizes the long-term reward while satisfying the constraints on the long-term cost. A …