ملف الباحث

Xiaoliang Hu

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. TVDO: Tchebycheff Value-Decomposition Optimization for Multiagent Reinforcement Learning

    2024 · IEEE Transactions on Neural Networks and Learning Systems

    In cooperative multiagent reinforcement learning (MARL), centralized training with decentralized execution (CTDE) has recently attracted more attention due to the physical demand. However, the most dilemma therein is the inconsistency between jointly-trained policies and individually …