DALL·E 2 brings photorealistic text-to-image generation
OpenAI's DALL·E 2 used a diffusion decoder conditioned on CLIP embeddings to generate high-resolution, photorealistic images from text, kicking off 2022's image-generation boom alongside Midjourney and Stable Diffusion.
Key facts
- Announced 6 April 2022
- Paper: 'Hierarchical Text-Conditional Image Generation with CLIP Latents', arXiv 2204.06125
- Supported inpainting and image variations
- Opened to the public without a waitlist in September 2022
What happened
DALL·E 2 produced images of a quality that made text-to-image a mainstream phenomenon.
Why it matters
Marked diffusion models' takeover of image generation and triggered debates on artists' rights and synthetic media.
Changelog
- 2026-09-29: created
Related events
Sources (2)
- officialDALL·E 2 (OpenAI)
- paperHierarchical Text-Conditional Image Generation with CLIP Latents (arXiv)
id: 2022-04-06-dall-e-2 · updated 2026-09-29 · open in the interactive timeline