ملف الباحث

Yara Rizk

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls

    2024 · arXiv (Cornell University)

    The resurgence of autonomous agents built using large language models (LLMs) to solve complex real-world tasks has brought increased focus on LLMs' fundamental ability of tool or function calling. At the core of these agents, …