Researcher profile

Yadong Li

2 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. TVDO: Tchebycheff Value-Decomposition Optimization for Multiagent Reinforcement Learning

    2024 · IEEE Transactions on Neural Networks and Learning Systems

    In cooperative multiagent reinforcement learning (MARL), centralized training with decentralized execution (CTDE) has recently attracted more attention due to the physical demand. However, the most dilemma therein is the inconsistency between jointly-trained policies and individually …

  2. NGRPO: Negative-enhanced Group Relative Policy Optimization

    2025 · arXiv (Cornell University)

    RLVR has enhanced the reasoning capabilities of Large Language Models (LLMs) across various tasks. However, GRPO, a representative RLVR algorithm, suffers from a critical limitation: when all responses within a group are either entirely correct …