ملف الباحث
Tom Lefebvre
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Information-Theoretic Policy Learning from Partial Observations with Fully Informed Decision Makers
2022 · arXiv (Cornell University)
In this work we formulate and treat an extension of the Imitation from Observations problem. Imitation from Observations is a generalisation of the well-known Imitation Learning problem where state-only demonstrations are considered. In our treatment …
-
A New Strategy for Incorporating Gaussian Process Dynamic Models into Stochastic Dynamic Programming
2025
This paper proposes a local solution method tailored to stochastic optimal control problems with Gaussian process (GP) representation of the dynamics leaning on the stochastic dynamic programming (DP) approach. We explore two methods—Fourier-Hermite DP (FHDP) …