ملف الباحث
Bingxin Zhou
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
LLQL: Logistic Likelihood Q-Learning for Reinforcement Learning
2023 · arXiv (Cornell University)
Modern reinforcement learning (RL) can be categorized into online and offline variants. As a pivotal aspect of both online and offline RL, current research on the Bellman equation revolves primarily around optimization techniques and performance …