ملف الباحث
Ruilin Luo
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
2024
The ability of Large Language Models (LLMs) to critique and refine their reasoning is crucial for their application in evaluation, feedback provision, and self-improvement.This paper introduces CRITICBENCH, a comprehensive benchmark designed to assess LLMs' abilities …