ملف الباحث
Kiritchenko, Svetlana
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Fine-Tuning Lowers Safety and Disrupts Evaluation Consistency
2025 · arXiv (Cornell University)
Fine-tuning a general-purpose large language model (LLM) for a specific domain or task has become a routine procedure for ordinary users. However, fine-tuning is known to remove the safety alignment features of the model, even …