ملف الباحث
Kostas G. Papakonstantinou
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Convex Is Back: Solving Belief MDPs With Convexity-Informed Deep Reinforcement Learning
2025 · arXiv (Cornell University)
We present a novel method for Deep Reinforcement Learning (DRL), incorporating the convex property of the value function over the belief space in Partially Observable Markov Decision Processes (POMDPs). We introduce hard- and soft-enforced convexity …