ملف الباحث
Chandler Zhou
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Aligning Language Models with Offline Learning from Human Feedback
2023 · arXiv (Cornell University)
Learning from human preferences is crucial for language models (LMs) to effectively cater to human needs and societal values. Previous research has made notable progress by leveraging human feedback to follow instructions. However, these approaches …