ملف الباحث
Rahul Nair
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Iterative Reward Shaping using Human Feedback for Correcting Reward Misspecification
2023 · arXiv (Cornell University)
A well-defined reward function is crucial for successful training of an reinforcement learning (RL) agent. However, defining a suitable reward function is a notoriously challenging task, especially in complex, multi-objective environments. Developers often have to …