ملف الباحث
Thomas Spooner
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
A Natural Actor-Critic Algorithm with Downside Risk Constraints
2020 · arXiv (Cornell University)
Existing work on risk-sensitive reinforcement learning - both for symmetric and downside risk measures - has typically used direct Monte-Carlo estimation of policy gradients. While this approach yields unbiased gradient estimates, it also suffers from …