DALL·E 3
Rewrites a long prompt into a detailed description, then draws the image
WHAT IT IS
DALL·E 3 is OpenAI’s text-to-image model, released in October 2023. It folds prompt rewriting and image generation into one flow: the model first expands a short user prompt into a more detailed scene description, then generates the image from it, which improves adherence to long prompts and complex constraints. Training used automatic recaptioning to add fine-grained descriptions to images, narrowing the gap between images and their text labels. It is available through the API and ChatGPT.
Why it matters
It made “understand the prompt first, then draw” a single pipeline, markedly improving adherence to long prompts and constraints, and turned conversational image generation into a routine ChatGPT capability.
Key specs
- Resolution
- 1024×1024, 1792×1024, 1024×1792
- Prompt handling
- Automatic recaptioning during training
- Open weights
- No
- Availability
- API and ChatGPT
Capabilities
Related concepts
Diffusion Models
Learn a thousand tiny denoising steps, and you can build an image from pure noise
Latent Diffusion & Conditional Control
Run diffusion not over pixels, but inside a compressed semantic space
Multimodal Generation
One model that learns to speak, to draw, to move — even to model the 3D world
Prompting & Alignment
Making a model helpful, honest and harmless is harder than simply making it bigger
Comparable products
Midjourney
2022A text-to-image service known for its aesthetic style
Stable Diffusion
2022Released text-to-image weights openly and small enough to run on consumer GPUs
FLUX
2024Generates high-resolution images with a rectified-flow transformer
Imagen
2022Generates photorealistic images with cascaded diffusion and a large text encoder
Firefly
2023An image generation and editing tool aimed at creators
Seedream
2024A text-to-image model with native high resolution and strong text rendering