Researcher profile

Andreas Terzis

1 paper in the PaperMetrix corpus

Publications

Papers by this author

  1. The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections

    2025 · arXiv (Cornell University)

    How should we evaluate the robustness of language model defenses? Current defenses against jailbreaks and prompt injections (which aim to prevent an attacker from eliciting harmful knowledge or remotely triggering malicious actions, respectively) are typically …