2023

AltDiffusion: A Multilingual Text-to-Image Diffusion Model

Ye, Fulong, Liu, Guang, Wu, Xinya et al.

Understand

Large Text-to-Image(T2I) diffusion models have shown a remarkable capability to produce photorealistic and diverse images based on text inputs.

  • However, existing works only support limited language input, e.g., English, Chinese, and Japanese, leaving users beyond these languages underserved and blocking the global expansion of T2I models.
  • Therefore, this paper presents AltDiffusion, a novel multilingual T2I diffusion model that supports eighteen different languages.
  • Specifically, we first train a multilingual text encoder based on the knowledge distillation.

Reading the bibliography…