ملف الباحث

Junfeng Fang

ورقتان في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. On the Role of Attention Heads in Large Language Model Safety

    2024 · arXiv (Cornell University)

    Large language models (LLMs) achieve state-of-the-art performance on multiple language tasks, yet their safety guardrails can be circumvented, leading to harmful generations. In light of this, recent research on safety mechanisms has emerged, revealing that …

  2. Reinforced Lifelong Editing for Language Models

    2025 · arXiv (Cornell University)

    Large language models (LLMs) acquire information from pre-training corpora, but their stored knowledge can become inaccurate or outdated over time. Model editing addresses this challenge by modifying model parameters without retraining, and prevalent approaches leverage …