ملف الباحث

Ali Mousavi

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Off-policy Evaluation in Infinite-Horizon Reinforcement Learning with Latent Confounders

    2020 · arXiv (Cornell University)

    Off-policy evaluation (OPE) in reinforcement learning is an important problem in settings where experimentation is limited, such as education and healthcare. But, in these very same settings, observed actions are often confounded by unobserved variables …