conference-paper

Optimizing Reinforcement Learning in Partially Observable Environments Using Compressed Suffix Memory Algorithm

Research footprint

At a glance

Citations
8
References
10
Comments
0
Paper overview

Abstract

Reinforcement learning in partially observable environments poses challenges due to limited and noisy observations. Traditional approaches like the Utile Suffix Memory (USM) algorithm suffer from inefficiencies and potential overfitting. In this paper, I propose the Compressed Suffix Memory (CSM) algorithm, designed to enhance state space generation and decision-making efficiency. CSM leverages heuristic information obtained from initial blind exploration of the environment to dynamically adjust tree depth and instance density thresholds. By incorporating Boltzmann sampling, CSM balances exploration and exploitation, thereby improving learning performance. Experimental results on benchmark mazes demonstrate that CSM outperforms USM in terms of learning speed and effectiveness, providing a promising advancement in reinforcement learning algorithms for complex, partially observable domains.

Record transparency

Publication details

DOI
10.1109/iceace63551.2024.10899025
OpenAlex
W4408100852
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.