ملف الباحث

An-An Liu

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Reason2Attack: Jailbreaking Text-to-Image Models via LLM Reasoning

    2025 · arXiv (Cornell University)

    Text-to-Image(T2I) models typically deploy safety filters to prevent the generation of sensitive images. Unfortunately, recent jailbreaking attack methods manually design instructions for the LLM to generate adversarial prompts, which effectively bypass safety filters while producing …