ملف الباحث
Hye-Bin Shin
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Aligning Humans and Robots via Reinforcement Learning from Implicit Human Feedback
2025 · arXiv (Cornell University)
Conventional reinforcement learning (RL) ap proaches often struggle to learn effective policies under sparse reward conditions, necessitating the manual design of complex, task-specific reward functions. To address this limitation, rein forcement learning from human feedback …