Researcher profile

Zhenghao Peng

2 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. Meta Proximal Policy Optimization for Cooperative Multi-Agent Continuous Control

    2022 · 2022 International Joint Conference on Neural Networks (IJCNN)

    In this paper we propose Multi-Agent Proxy Proximal Policy Optimization (MA3PO), a novel multi-agent deep reinforcement learning algorithm that tackles the challenge of cooperative continuous multi-agent control. Our method is driven by the observation that …

  2. Guarded Policy Optimization with Imperfect Online Demonstrations

    2023 · arXiv (Cornell University)

    The Teacher-Student Framework (TSF) is a reinforcement learning setting where a teacher agent guards the training of a student agent by intervening and providing online demonstrations. Assuming optimal, the teacher policy has the perfect timing …