Researcher profile

Lijun Ding

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Provably Convergent Policy Optimization via Metric-aware Trust Region Methods

    2023 · arXiv (Cornell University)

    Trust-region methods based on Kullback-Leibler divergence are pervasively used to stabilize policy optimization in reinforcement learning. In this paper, we exploit more flexible metrics and examine two natural extensions of policy optimization with Wasserstein and …