ملف الباحث

Bai, Weiheng

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Evolving Security in LLMs: A Study of Jailbreak Attacks and Defenses

    2025 · arXiv (Cornell University)

    Large Language Models (LLMs) are increasingly popular, powering a wide range of applications. Their widespread use has sparked concerns, especially through jailbreak attacks that bypass safety measures to produce harmful content. In this paper, we …