ملف الباحث

Zhenhong Zhou

4 أوراق في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. On the Role of Attention Heads in Large Language Model Safety

    2024 · arXiv (Cornell University)

    Large language models (LLMs) achieve state-of-the-art performance on multiple language tasks, yet their safety guardrails can be circumvented, leading to harmful generations. In light of this, recent research on safety mechanisms has emerged, revealing that …

  2. Reinforced Lifelong Editing for Language Models

    2025 · arXiv (Cornell University)

    Large language models (LLMs) acquire information from pre-training corpora, but their stored knowledge can become inaccurate or outdated over time. Model editing addresses this challenge by modifying model parameters without retraining, and prevalent approaches leverage …

  3. A Vision for Auto Research with LLM Agents

    2025 · arXiv (Cornell University)

    This paper introduces Agent-Based Auto Research, a structured multi-agent framework designed to automate, coordinate, and optimize the full lifecycle of scientific research. Leveraging the capabilities of large language models (LLMs) and modular agent collaboration, the …

  4. Hidden in the Noise: Unveiling Backdoors in Audio LLMs Alignment Through Latent Acoustic Pattern Triggers

    2026 · Proceedings of the AAAI Conference on Artificial Intelligence

    As Audio Large Language Models (ALLMs) emerge as powerful tools for speech processing, their safety implications demand urgent attention. While considerable research has explored textual and vision safety, audio’s distinct characteristics present significant challenges. This …