ملف الباحث
Wei, Wenlan
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Evolving Security in LLMs: A Study of Jailbreak Attacks and Defenses
2025 · arXiv (Cornell University)
Large Language Models (LLMs) are increasingly popular, powering a wide range of applications. Their widespread use has sparked concerns, especially through jailbreak attacks that bypass safety measures to produce harmful content. In this paper, we …