ملف الباحث
Jeffery M. Capone
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Evaluating the Impact of Compression Techniques on Task-Specific Performance of Large Language Models
2024 · arXiv (Cornell University)
Large language models (LLMs) offer powerful capabilities but incur substantial computational costs, driving the need for efficient compression techniques. This study evaluates the impact of popular compression methods - Magnitude Pruning, SparseGPT, and Wanda - …