ملف الباحث
Lucas Weber
ورقتان في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
The ICL Consistency Test
2023 · arXiv (Cornell University)
Just like the previous generation of task-tuned models, large language models (LLMs) that are adapted to tasks via prompt-based methods like in-context-learning (ICL) perform well in some setups but not in others. This lack of …
-
tinyBenchmarks: evaluating LLMs with fewer examples
2024 · arXiv (Cornell University)
The versatility of large language models (LLMs) led to the creation of diverse benchmarks that thoroughly test a variety of language models' abilities. These benchmarks consist of tens of thousands of examples making evaluation of …