ملف الباحث
Joseph Modayil
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Adapting the Function Approximation Architecture in Online Reinforcement Learning
2021 · arXiv (Cornell University)
The performance of a reinforcement learning (RL) system depends on the computational architecture used to approximate a value function. Deep learning methods provide both optimization techniques and architectures for approximating nonlinear functions from noisy, high-dimensional …