ملف الباحث
Li Ping
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Shaping in reinforcement learning via knowledge transferred from human-demonstrations
2015
Transfer has been widely used to ameliorate the slow convergence speed of reinforcement learning (RL) by reusing the previous obtained knowledge from other related but distinct tasks. In this paper, we propose a framework to …
-
AIBox
2019
As one of the major search engines in the world, Baidu's Sponsored Search has long adopted the use of deep neural network (DNN) models for Ads click-through rate (CTR) predictions, as early as in 2013. …