ملف الباحث
An-An Liu
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Reason2Attack: Jailbreaking Text-to-Image Models via LLM Reasoning
2025 · arXiv (Cornell University)
Text-to-Image(T2I) models typically deploy safety filters to prevent the generation of sensitive images. Unfortunately, recent jailbreaking attack methods manually design instructions for the LLM to generate adversarial prompts, which effectively bypass safety filters while producing …