Researcher profile
Thiago D. Simão
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Structure Learning for Safe Policy Improvement
2019
We investigate how Safe Policy Improvement (SPI) algorithms can exploit the structure of factored Markov decision processes when such structure is unknown a priori. To facilitate the application of reinforcement learning in the real world, …