ملف الباحث
Padmapriya Muthu
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Benchmarking clinical reasoning and accuracy of large language models on breast oncology multiple-choice questions.
2025 · Journal of Clinical Oncology
e13637 Background: Large language models (LLMs) like GPT-4 (OpenAI) and Claude Opus (Anthropic) showed high accuracy in medical multiple-choice exams, but data on their oncology-specific clinical reasoning and performance is limited. This study evaluates their …