ملف الباحث
Tom Hosking
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Human Feedback is not Gold Standard
2023 · arXiv (Cornell University)
Human feedback has become the de facto standard for evaluating the performance of Large Language Models, and is increasingly being used as a training objective. However, it is not clear which properties of a generated …