ملف الباحث
Huiling Qin
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Discriminator-Weighted Offline Imitation Learning from Suboptimal Demonstrations
2022 · arXiv (Cornell University)
We study the problem of offline Imitation Learning (IL) where an agent aims to learn an optimal expert behavior policy without additional online environment interactions. Instead, the agent is provided with a supplementary offline dataset …