Researcher profile
Wei, Wenlan
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Evolving Security in LLMs: A Study of Jailbreak Attacks and Defenses
2025 · arXiv (Cornell University)
Large Language Models (LLMs) are increasingly popular, powering a wide range of applications. Their widespread use has sparked concerns, especially through jailbreak attacks that bypass safety measures to produce harmful content. In this paper, we …