ملف الباحث
Huanqian Yan
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
Boosting Jailbreak Transferability for Large Language Models
2024 · arXiv (Cornell University)
Large language models have drawn significant attention to the challenge of safe alignment, especially regarding jailbreak attacks that circumvent security measures to produce harmful content. To address the limitations of existing methods like GCG, which …