ملف الباحث

Chandler Zhou

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Aligning Language Models with Offline Learning from Human Feedback

    2023 · arXiv (Cornell University)

    Learning from human preferences is crucial for language models (LMs) to effectively cater to human needs and societal values. Previous research has made notable progress by leveraging human feedback to follow instructions. However, these approaches …