Researcher profile

Erxin Yu

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. CoSafe: Evaluating Large Language Model Safety in Multi-Turn Dialogue Coreference

    2024 · arXiv (Cornell University)

    As large language models (LLMs) constantly evolve, ensuring their safety remains a critical research problem. Previous red-teaming approaches for LLM safety have primarily focused on single prompt attacks or goal hijacking. To the best of …