Multilingual Conceptual Coverage in Text-to-Image Models
Michael Saxon, William Yang Wang · 2023
DallE mega1 Figure 1: A selection of images generated by DALLE-mega, Stable Diffusion 2, DALLE-2, and AltDiffusion, illustrating their conceptual coverage of "dog," "airplane," and "face" across English, Spanish, German, Chinese (simplified), Japanese, Hebrew, and Indonesian.Coverage of the concepts varies considerably across model and language, and can be observed in the consistency and correctness of images generated under simple prompts.