Researcher profile

Fanyu Que

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Accelerated Primal-Dual Policy Optimization for Safe Reinforcement Learning

    2018 · arXiv (Cornell University)

    Constrained Markov Decision Process (CMDP) is a natural framework for reinforcement learning tasks with safety constraints, where agents learn a policy that maximizes the long-term reward while satisfying the constraints on the long-term cost. A …