Researcher profile
Junxiao Yang
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints
2025 · arXiv (Cornell University)
Jailbreaking attacks can effectively induce unsafe behaviors in Large Language Models (LLMs); however, the transferability of these attacks across different models remains limited. This study aims to understand and enhance the transferability of gradient-based jailbreaking …