Researcher profile
Zhenghao Peng
2 papers in the PaperMetrix corpus
Publications
Papers by this author
-
Meta Proximal Policy Optimization for Cooperative Multi-Agent Continuous Control
2022 · 2022 International Joint Conference on Neural Networks (IJCNN)
In this paper we propose Multi-Agent Proxy Proximal Policy Optimization (MA3PO), a novel multi-agent deep reinforcement learning algorithm that tackles the challenge of cooperative continuous multi-agent control. Our method is driven by the observation that …
-
Guarded Policy Optimization with Imperfect Online Demonstrations
2023 · arXiv (Cornell University)
The Teacher-Student Framework (TSF) is a reinforcement learning setting where a teacher agent guards the training of a student agent by intervening and providing online demonstrations. Assuming optimal, the teacher policy has the perfect timing …