Researcher profile
Dawkins, Hillary
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Fine-Tuning Lowers Safety and Disrupts Evaluation Consistency
2025 · arXiv (Cornell University)
Fine-tuning a general-purpose large language model (LLM) for a specific domain or task has become a routine procedure for ordinary users. However, fine-tuning is known to remove the safety alignment features of the model, even …