ملف الباحث

Chawin Sitawarin

3 أوراق في مجموعة PaperMetrix

المنشورات

أوراق هذا المؤلف

  1. Demystifying the Adversarial Robustness of Random Transformation Defenses

    2022 · arXiv (Cornell University)

    Neural networks' lack of robustness against attacks raises concerns in security-sensitive settings such as autonomous vehicles. While many countermeasures may look promising, only a few withstand rigorous evaluation. Defenses using random transformations (RT) have shown …

  2. OODRobustBench: a Benchmark and Large-Scale Analysis of Adversarial Robustness under Distribution Shift

    2023 · arXiv (Cornell University)

    Existing works have made great progress in improving adversarial robustness, but typically test their method only on data from the same distribution as the training data, i.e. in-distribution (ID) testing. As a result, it is …

  3. The Attacker Moves Second: Stronger Adaptive Attacks Bypass Defenses Against Llm Jailbreaks and Prompt Injections

    2025 · arXiv (Cornell University)

    How should we evaluate the robustness of language model defenses? Current defenses against jailbreaks and prompt injections (which aim to prevent an attacker from eliciting harmful knowledge or remotely triggering malicious actions, respectively) are typically …