Researcher profile
Junfeng Fang
2 papers in the PaperMetrix corpus
Publications
Papers by this author
-
On the Role of Attention Heads in Large Language Model Safety
2024 · arXiv (Cornell University)
Large language models (LLMs) achieve state-of-the-art performance on multiple language tasks, yet their safety guardrails can be circumvented, leading to harmful generations. In light of this, recent research on safety mechanisms has emerged, revealing that …
-
Reinforced Lifelong Editing for Language Models
2025 · arXiv (Cornell University)
Large language models (LLMs) acquire information from pre-training corpora, but their stored knowledge can become inaccurate or outdated over time. Model editing addresses this challenge by modifying model parameters without retraining, and prevalent approaches leverage …