Researcher profile

Huang, MinLie

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak

    2024 · arXiv (Cornell University)

    Large Language Models (LLMs) are susceptible to generating harmful content when prompted with carefully crafted inputs, a vulnerability known as LLM jailbreaking. As LLMs become more powerful, studying jailbreak methods is critical to enhancing security …