ملف الباحث

Outongyi Lv

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. LLQL: Logistic Likelihood Q-Learning for Reinforcement Learning

    2023 · arXiv (Cornell University)

    Modern reinforcement learning (RL) can be categorized into online and offline variants. As a pivotal aspect of both online and offline RL, current research on the Bellman equation revolves primarily around optimization techniques and performance …