ملف الباحث

Murtaza Dalal

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. AWAC: Accelerating Online Reinforcement Learning with Offline Datasets

    2020 · arXiv (Cornell University)

    Reinforcement learning (RL) provides an appealing formalism for learning control policies from experience. However, the classic active formulation of RL necessitates a lengthy active exploration process for each behavior, making it difficult to apply in …