ملف الباحث

Thomas Spooner

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. A Natural Actor-Critic Algorithm with Downside Risk Constraints

    2020 · arXiv (Cornell University)

    Existing work on risk-sensitive reinforcement learning - both for symmetric and downside risk measures - has typically used direct Monte-Carlo estimation of policy gradients. While this approach yields unbiased gradient estimates, it also suffers from …