Researcher profile

Jinqi Luo

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. KDA: A Knowledge-Distilled Attacker for Generating Diverse Prompts to Jailbreak LLMs

    2025 · arXiv (Cornell University)

    Jailbreak attacks exploit specific prompts to bypass LLM safeguards, causing the LLM to generate harmful, inappropriate, and misaligned content. Current jailbreaking methods rely heavily on carefully designed system prompts and numerous queries to achieve a …