ملف الباحث
Jiajun Fan
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Critic PI2: Master Continuous Planning via Policy Improvement with Path Integrals and Deep Actor-Critic Reinforcement Learning
2020 · arXiv (Cornell University)
Constructing agents with planning capabilities has long been one of the main challenges in the pursuit of artificial intelligence. Tree-based planning methods from AlphaGo to Muzero have enjoyed huge success in discrete domains, such as …
-
Generalized Data Distribution Iteration
2022 · arXiv (Cornell University)
To obtain higher sample efficiency and superior final performance simultaneously has been one of the major challenges for deep reinforcement learning (DRL). Previous work could handle one of these challenges but typically failed to address …