ملف الباحث

Kriti Aggarwal

ورقة واحدة في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. RED QUEEN: Safeguarding Large Language Models against Concealed Multi-Turn Jailbreaking

    2024 · arXiv (Cornell University)

    The rapid progress of Large Language Models (LLMs) has opened up new opportunities across various domains and applications; yet it also presents challenges related to potential misuse. To mitigate such risks, red teaming has been …