Researcher profile
Erxin Yu
1 paper in the PaperMetrix corpus
Publications
Papers by this author
-
CoSafe: Evaluating Large Language Model Safety in Multi-Turn Dialogue Coreference
2024 · arXiv (Cornell University)
As large language models (LLMs) constantly evolve, ensuring their safety remains a critical research problem. Previous red-teaming approaches for LLM safety have primarily focused on single prompt attacks or goal hijacking. To the best of …