conference-paper Open access

On the Feasibility of Poisoning Text-to-Image AI Models via Adversarial Mislabeling

Research footprint

At a glance

Citations
1
References
15
Comments
0
Paper overview

Abstract

Today's text-to-image generative models are trained on millions of images sourced from the Internet, each paired with a detailed caption produced by Vision-Language Models (VLMs). This part of the training pipeline is critical for supplying the models with large volumes of high-quality image-caption pairs during training. However, recent work suggests that VLMs are vulnerable to stealthy adversarial attacks, where adversarial perturbations are added to images to mislead the VLMs into producing incorrect captions.

Record transparency

Publication details

DOI
10.1145/3719027.3744845
OpenAlex
W4416549620
Document type
conference-paper
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.