Researcher profile
Yadong Li
2 papers in the PaperMetrix corpus
Publications
Papers by this author
-
TVDO: Tchebycheff Value-Decomposition Optimization for Multiagent Reinforcement Learning
2024 · IEEE Transactions on Neural Networks and Learning Systems
In cooperative multiagent reinforcement learning (MARL), centralized training with decentralized execution (CTDE) has recently attracted more attention due to the physical demand. However, the most dilemma therein is the inconsistency between jointly-trained policies and individually …
-
NGRPO: Negative-enhanced Group Relative Policy Optimization
2025 · arXiv (Cornell University)
RLVR has enhanced the reasoning capabilities of Large Language Models (LLMs) across various tasks. However, GRPO, a representative RLVR algorithm, suffers from a critical limitation: when all responses within a group are either entirely correct …