ملف الباحث

C Hirayama

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Learning Stabilization Control from Observations by Learning Lyapunov-like Proxy Models

    2023

    The deployment of Reinforcement Learning to robotics applications faces the difficulty of reward engineering. Therefore, approaches have focused on creating reward functions by Learning from Observations (LfO) which is the task of learning policies from …