ملف الباحث
Huang, MinLie
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak
2024 · arXiv (Cornell University)
Large Language Models (LLMs) are susceptible to generating harmful content when prompted with carefully crafted inputs, a vulnerability known as LLM jailbreaking. As LLMs become more powerful, studying jailbreak methods is critical to enhancing security …