ملف الباحث

Zhu, Junda

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak

    2024 · arXiv (Cornell University)

    Large Language Models (LLMs) are susceptible to generating harmful content when prompted with carefully crafted inputs, a vulnerability known as LLM jailbreaking. As LLMs become more powerful, studying jailbreak methods is critical to enhancing security …