ملف الباحث

Shicong Cen

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Fast Global Convergence of Natural Policy Gradient Methods with Entropy Regularization

    2020 · arXiv (Cornell University)

    Natural policy gradient (NPG) methods are among the most widely used policy optimization algorithms in contemporary reinforcement learning. This class of methods is often applied in conjunction with entropy regularization -- an algorithmic scheme that …