article Open access

Synthetic Experiences for Accelerating DQN Performance in Discrete Non-Deterministic Environments

  • Algorithms
  • Multidisciplinary Digital Publishing Institute
Research footprint

At a glance

Citations
4
References
29
Comments
0
Paper overview

Abstract

State-of-the-art Deep Reinforcement Learning Algorithms such as DQN and DDPG use the concept of a replay buffer called Experience Replay. The default usage contains only the experiences that have been gathered over the runtime. We propose a method called Interpolated Experience Replay that uses stored (real) transitions to create synthetic ones to assist the learner. In this first approach to this field, we limit ourselves to discrete and non-deterministic environments and use a simple equally weighted average of the reward in combination with observed follow-up states. We could demonstrate a significantly improved overall mean average in comparison to a DQN network with vanilla Experience Replay on the discrete and non-deterministic FrozenLake8x8-v0 environment.

Record transparency

Publication details

DOI
10.3390/a14080226
OpenAlex
W3184593887
Document type
article
Language
EN
Source
Algorithms
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.