ملف الباحث
Sakyasingha Dasgupta
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Internal Model from Observations for Reward Shaping
2018 · arXiv (Cornell University)
Reinforcement learning methods require careful design involving a reward function to obtain the desired action policy for a given task. In the absence of hand-crafted reward functions, prior work on the topic has proposed several …
-
Continual Learning via Online Leverage Score Sampling
2019 · arXiv (Cornell University)
In order to mimic the human ability of continual acquisition and transfer of knowledge across various tasks, a learning system needs the capability for continual learning, effectively utilizing the previously acquired skills. As such, the …