ملف الباحث

Steven Hansen

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Fast deep reinforcement learning using online adjustments from the past

    2018 · arXiv (Cornell University)

    We propose Ephemeral Value Adjusments (EVA): a means of allowing deep reinforcement learning agents to rapidly adapt to experience in their replay buffer. EVA shifts the value predicted by a neural network with an estimate …