ملف الباحث
Yuwei Han
ورقة واحدة في مجموعة PaperMetrix
المنشورات
أوراق هذا المؤلف
-
AdvEvo-MARL: Shaping Internalized Safety through Adversarial Co-Evolution in Multi-Agent Reinforcement Learning
2025 · arXiv (Cornell University)
LLM-based multi-agent systems excel at planning, tool use, and role coordination, but their openness and interaction complexity also expose them to jailbreak, prompt-injection, and adversarial collaboration. Existing defenses fall into two lines: (i) self-verification that …