ملف الباحث

Esraa Elelimy

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Deep Reinforcement Learning with Gradient Eligibility Traces

    2025 · arXiv (Cornell University)

    Achieving fast and stable off-policy learning in deep reinforcement learning (RL) is challenging. Most existing methods rely on semi-gradient temporal-difference (TD) methods for their simplicity and efficiency, but are consequently susceptible to divergence. While more …