ملف الباحث
Andrew Bennett
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Off-policy Evaluation in Infinite-Horizon Reinforcement Learning with Latent Confounders
2020 · arXiv (Cornell University)
Off-policy evaluation (OPE) in reinforcement learning is an important problem in settings where experimentation is limited, such as education and healthcare. But, in these very same settings, observed actions are often confounded by unobserved variables …