The Biased Artist: Exploiting Cultural Biases via Homoglyphs in Text-Guided Image Generation Models

There is no author summary for this article yet. Authors can add summaries to their articles on ScienceOpen to make them more accessible to a non-specialist audience.

Abstract

Text-guided image generation models, such as DALL-E 2 and Stable Diffusion, have recently received much attention from academia and the general public. Provided with textual descriptions, these models are capable of generating high-quality images depicting various concepts and styles. However, such models are trained on large amounts of public data and implicitly learn relationships from their training data that are not immediately apparent. We demonstrate that common multimodal models implicitly learned cultural biases that can be triggered and injected into the generated images by simply replacing single characters in the textual description with visually similar non-Latin characters. These so-called homoglyph replacements enable malicious users or service providers to induce biases into the generated images and even render the whole generation process useless. We practically illustrate such attacks on DALL-E 2 and Stable Diffusion as text-guided image generation models and further show that CLIP also behaves similarly. Our results further indicate that text encoders trained on multilingual data provide a way to mitigate the effects of homoglyph replacements.

Related collections

Author and article information

Journal

Publication date Created: 19 September 2022

Article

ArXiV ID: 2209.08891

SO-VID: 222a5977-82bb-4849-8742-b1c796e877f9

License:

http://arxiv.org/licenses/nonexclusive-distrib/1.0/

History

Custom metadata

Comments 31 pages, 19 figures, 4 tables

Categories cs.CV cs.AI cs.CY cs.LG

ScienceOpen disciplines: Computer vision & Pattern recognition,Applied computer science,Artificial intelligence

Data availability:

ScienceOpen disciplines: Computer vision & Pattern recognition, Applied computer science, Artificial intelligence

The Biased Artist: Exploiting Cultural Biases via Homoglyphs in Text-Guided Image Generation Models

Read this article at

Abstract

Related collections

Radiology and Natural Language Processing

Author and article information

Journal

Article

History

Custom metadata

Comments

Comment on this article

Similar content 93