ملف الباحث
Yara Rizk
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
NESTFUL: A Benchmark for Evaluating LLMs on Nested Sequences of API Calls
2024 · arXiv (Cornell University)
The resurgence of autonomous agents built using large language models (LLMs) to solve complex real-world tasks has brought increased focus on LLMs' fundamental ability of tool or function calling. At the core of these agents, …