ملف الباحث

Karim Beguir

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Ranked Reward: Enabling Self-Play Reinforcement Learning for Combinatorial Optimization

    2018 · arXiv (Cornell University)

    Adversarial self-play in two-player games has delivered impressive results when used with reinforcement learning algorithms that combine deep neural networks and tree search. Algorithms like AlphaZero and Expert Iteration learn tabula-rasa, producing highly informative training …

  2. Fast Population-Based Reinforcement Learning on a Single Machine

    2022 · arXiv (Cornell University)

    Training populations of agents has demonstrated great promise in Reinforcement Learning for stabilizing training, improving exploration and asymptotic performance, and generating a diverse set of solutions. However, population-based training is often not considered by practitioners …