ملف الباحث
Aihua Pei
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
SelfPrompt: Autonomously Evaluating LLM Robustness via Domain-Constrained Knowledge Guidelines and Refined Adversarial Prompts
2024 · arXiv (Cornell University)
Traditional methods for evaluating the robustness of large language models (LLMs) often rely on standardized benchmarks, which can escalate costs and limit evaluations across varied domains. This paper introduces a novel framework designed to autonomously …