arXiv · 2303.07909
Text-to-image Diffusion Models in Generative AI: A Survey
Abstract
This survey reviews the progress of diffusion models in generating images from text, ~\textit{i.e.} text-to-image diffusion models. As a self-contained work, this survey starts with a brief introduction of how diffusion models work for image synthesis, followed by the background for text-conditioned image synthesis. Based on that, we present an organized review of pioneering methods and their improvements on text-to-image generation. We further summarize applications beyond image generation, such as text-guided generation for various modalities like videos, and text-guided image editing. Beyond the progress made so far, we discuss existing challenges and promising future directions.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Chenshuang Zhang, Chaoning Zhang, Mengchun Zhang, In So Kweon, Junmo Kim. 2023-03-14. Text-to-image Diffusion Models in Generative AI: A Survey. https://arxiv.org/abs/2303.07909
Cite the original work for its findings. Save a collection to share your selection of sources.