Researcher profile

Hepeng Li

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. An Improved Trust-Region Method for Off-Policy Deep Reinforcement Learning

    2023

    Reinforcement learning (RL) is a powerful tool for training agents to interact with complex environments. In particular, trust-region methods are widely used for policy optimization in model-free RL. However, these methods suffer from high sample …