ملف الباحث
Yihuan Mao
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
MOORe: Model-based Offline-to-Online Reinforcement Learning
2022 · arXiv (Cornell University)
With the success of offline reinforcement learning (RL), offline trained RL policies have the potential to be further improved when deployed online. A smooth transfer of the policy matters in safe real-world deployment. Besides, fast …