ملف الباحث

Wenqi Cai

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Quasi-Newton Iteration in Deterministic Policy Gradient

    2022 · arXiv (Cornell University)

    This paper presents a model-free approximation for the Hessian of the performance of deterministic policies to use in the context of Reinforcement Learning based on Quasi-Newton steps in the policy parameters. We show that the …