ملف الباحث
Shuyi Liu
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
JailBench: A Comprehensive Chinese Security Assessment Benchmark for Large Language Models
2025 · arXiv (Cornell University)
Large language models (LLMs) have demonstrated remarkable capabilities across various applications, highlighting the urgent need for comprehensive safety evaluations. In particular, the enhanced Chinese language proficiency of LLMs, combined with the unique characteristics and complexity …