conference-paper

Modified Annealed Adversarial Bonus for Adversarially Guided Actor-Critic

  • 2022 37th Youth Academic Annual Conference of Chinese Association of Automation (YAC)
Research footprint

At a glance

Citations
0
References
22
Comments
0
Paper overview

Abstract

This paper investigates learning efficiency for rein-forcement learning in procedurally generated environments. A more sophisticated method is proposed to adjust the adversarial bonus to promote learning efficiency instead of the linearly decayed scheme in adversarially guided actor-critic. Our method considers the relationship between the bonus adjustment and the learning procedure. In some environments, if an agent performs better in learning, the agent will reach the goal with fewer steps. If the length of the episode decreases, the adversarial bonus will be reduced in our method. In this way, the learning efficiency has been improved in some procedurally generated tasks. Several experiments are implemented in MiniGrid to verify the proposed method. In the experiments, the proposed method outperforms the existing adversarially guided methods in several challenging procedurally-generated tasks.

Record transparency

Publication details

DOI
10.1109/yac57282.2022.10023796
OpenAlex
W4318606477
Document type
conference-paper
Language
EN
Source
2022 37th Youth Academic Annual Conference of Chinese Association of Automation (YAC)
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.