Researcher profile

Bozhidar Vasilev

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Semi-On-Policy Training for Sample Efficient Multi-Agent Policy Gradients

    2021 · arXiv (Cornell University)

    Policy gradient methods are an attractive approach to multi-agent reinforcement learning problems due to their convergence properties and robustness in partially observable scenarios. However, there is a significant performance gap between state-of-the-art policy gradient and …