Researcher profile
X. Zhong
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
An Improved Trust-Region Method for Off-Policy Deep Reinforcement Learning
2023
Reinforcement learning (RL) is a powerful tool for training agents to interact with complex environments. In particular, trust-region methods are widely used for policy optimization in model-free RL. However, these methods suffer from high sample …