ملف الباحث
Yuting Wei
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Fast Global Convergence of Natural Policy Gradient Methods with Entropy Regularization
2020 · arXiv (Cornell University)
Natural policy gradient (NPG) methods are among the most widely used policy optimization algorithms in contemporary reinforcement learning. This class of methods is often applied in conjunction with entropy regularization -- an algorithmic scheme that …