Researcher profile
Huanqian Yan
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Boosting Jailbreak Transferability for Large Language Models
2024 · arXiv (Cornell University)
Large language models have drawn significant attention to the challenge of safe alignment, especially regarding jailbreak attacks that circumvent security measures to produce harmful content. To address the limitations of existing methods like GCG, which …