ملف الباحث
Tengyu Xu
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
When Will Generative Adversarial Imitation Learning Algorithms Attain Global Convergence
2020 · arXiv (Cornell University)
Generative adversarial imitation learning (GAIL) is a popular inverse reinforcement learning approach for jointly optimizing policy and reward from expert trajectories. A primary question about GAIL is whether applying a certain policy gradient algorithm to …
-
Improving Sample Complexity Bounds for (Natural) Actor-Critic Algorithms
2020 · Neural Information Processing Systems
The actor-critic (AC) algorithm is a popular method to find an optimal policy in reinforcement learning. In the infinite horizon scenario, the finite-sample convergence rate for the AC and natural actor-critic (NAC) algorithms has been …