ملف الباحث
Erxin Yu
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
CoSafe: Evaluating Large Language Model Safety in Multi-Turn Dialogue Coreference
2024 · arXiv (Cornell University)
As large language models (LLMs) constantly evolve, ensuring their safety remains a critical research problem. Previous red-teaming approaches for LLM safety have primarily focused on single prompt attacks or goal hijacking. To the best of …