ملف الباحث
Landon Kraemer
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Sample Bounded Distributed Reinforcement Learning for Decentralized POMDPs
2021 · Proceedings of the AAAI Conference on Artificial Intelligence
Decentralized partially observable Markov decision processes (Dec-POMDPs) offer a powerful modeling technique for realistic multi-agent coordination problems under uncertainty. Prevalent solution techniques are centralized and assume prior knowledge of the model. We propose a distributed …