Researcher profile
Shiwen Cui
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
UpSafe$^\circ$C: Upcycling for Controllable Safety in Large Language Models
2025 · arXiv (Cornell University)
Large Language Models (LLMs) have achieved remarkable progress across a wide range of tasks, but remain vulnerable to safety risks such as harmful content generation and jailbreak attacks. Existing safety techniques -- including external guardrails, …