ملف الباحث

Zhiwei Shang

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Relative Entropy Regularized Sample Efficient Reinforcement Learning with Continuous Actions

    2022

    In this paper, a novel reinforcement learning (RL) approach, continuous dynamic policy programming (CDPP) is proposed to tackle the issues of both learning stability and sample efficiency in the current RL methods with continuous actions. …