ملف الباحث
Xiaoliang Hu
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
TVDO: Tchebycheff Value-Decomposition Optimization for Multiagent Reinforcement Learning
2024 · IEEE Transactions on Neural Networks and Learning Systems
In cooperative multiagent reinforcement learning (MARL), centralized training with decentralized execution (CTDE) has recently attracted more attention due to the physical demand. However, the most dilemma therein is the inconsistency between jointly-trained policies and individually …