ملف الباحث

Tengyu Xu

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. When Will Generative Adversarial Imitation Learning Algorithms Attain Global Convergence

    2020 · arXiv (Cornell University)

    Generative adversarial imitation learning (GAIL) is a popular inverse reinforcement learning approach for jointly optimizing policy and reward from expert trajectories. A primary question about GAIL is whether applying a certain policy gradient algorithm to …

  2. Improving Sample Complexity Bounds for (Natural) Actor-Critic Algorithms

    2020 · Neural Information Processing Systems

    The actor-critic (AC) algorithm is a popular method to find an optimal policy in reinforcement learning. In the infinite horizon scenario, the finite-sample convergence rate for the AC and natural actor-critic (NAC) algorithms has been …