ملف الباحث
Stephanie Lin
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
TruthfulQA: Measuring How Models Mimic Human Falsehoods
2021 · arXiv (Cornell University)
We propose a benchmark to measure whether a language model is truthful in generating answers to questions. The benchmark comprises 817 questions that span 38 categories, including health, law, finance and politics. We crafted questions …