ملف الباحث

Ritu Sidgal

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Benchmarking clinical reasoning and accuracy of large language models on breast oncology multiple-choice questions.

    2025 · Journal of Clinical Oncology

    e13637 Background: Large language models (LLMs) like GPT-4 (OpenAI) and Claude Opus (Anthropic) showed high accuracy in medical multiple-choice exams, but data on their oncology-specific clinical reasoning and performance is limited. This study evaluates their …