Researcher profile

Tsung-Yi Ho

3 papers in the PaperMetrix corpus

Publications

Papers by this author

  1. RADAR: Robust AI-Text Detection via Adversarial Learning

    2023 · arXiv (Cornell University)

    Recent advances in large language models (LLMs) and the intensifying popularity of ChatGPT-like applications have blurred the boundary of high-quality text generation between humans and machines. However, in addition to the anticipated revolutionary changes to …

  2. The Devil is in the Neurons: Interpreting and Mitigating Social Biases in Pre-trained Language Models

    2024 · arXiv (Cornell University)

    Pre-trained Language models (PLMs) have been acknowledged to contain harmful information, such as social biases, which may cause negative social impacts or even bring catastrophic results in application. Previous works on this problem mainly focused …

  3. Why LLM Safety Guardrails Collapse After Fine-tuning: A Similarity Analysis Between Alignment and Fine-tuning Datasets

    2026

    Lei Hsiung, Tianyu Pang, Yung-Chen Tang, Linyue Song, Tsung-Yi Ho, Pin-Yu Chen, Yaoqing Yang. Proceedings of the 64th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2026.