Researcher profile
Huang, MinLie
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak
2024 · arXiv (Cornell University)
Large Language Models (LLMs) are susceptible to generating harmful content when prompted with carefully crafted inputs, a vulnerability known as LLM jailbreaking. As LLMs become more powerful, studying jailbreak methods is critical to enhancing security …