ملف الباحث

Juan Hernandez-Garcia

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Multi-Step Reinforcement Learning: A Unifying Algorithm

    2018 · Proceedings of the AAAI Conference on Artificial Intelligence

    Unifying seemingly disparate algorithmic ideas to produce better performing algorithms has been a longstanding goal in reinforcement learning. As a primary example, TD(λ) elegantly unifies one-step TD prediction with Monte Carlo methods through the use …