Researcher profile

Thiago D. Simão

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. Structure Learning for Safe Policy Improvement

    2019

    We investigate how Safe Policy Improvement (SPI) algorithms can exploit the structure of factored Markov decision processes when such structure is unknown a priori. To facilitate the application of reinforcement learning in the real world, …