ملف الباحث
Kriti Aggarwal
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
RED QUEEN: Safeguarding Large Language Models against Concealed Multi-Turn Jailbreaking
2024 · arXiv (Cornell University)
The rapid progress of Large Language Models (LLMs) has opened up new opportunities across various domains and applications; yet it also presents challenges related to potential misuse. To mitigate such risks, red teaming has been …