ملف الباحث
Cameron Voloshin
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Empirical Study of Off-Policy Policy Evaluation for Reinforcement Learning
2019 · arXiv (Cornell University)
We offer an experimental benchmark and empirical study for off-policy policy evaluation (OPE) in reinforcement learning, which is a key problem in many safety critical applications. Given the increasing interest in deploying learning-based methods, there …