本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。
これは何か
DALL·E 3 is OpenAI’s text-to-image model, released in October 2023. It folds prompt rewriting and image generation into one flow: the model first expands a short user prompt into a more detailed scene description, then generates the image from it, which improves adherence to long prompts and complex constraints. Training used automatic recaptioning to add fine-grained descriptions to images, narrowing the gap between images and their text labels. It is available through the API and ChatGPT.
なぜ覚えておく価値があるか
It made “understand the prompt first, then draw” a single pipeline, markedly improving adherence to long prompts and constraints, and turned conversational image generation into a routine ChatGPT capability.
主な仕様
- Resolution
- 1024×1024, 1792×1024, 1024×1792
- Prompt handling
- Automatic recaptioning during training
- Open weights
- No
- Availability
- API and ChatGPT
対応する能力
関連する概念
同種の製品
Midjourney
2022美的スタイルで知られるテキストから画像生成サービス
Stable Diffusion
2022テキストから画像生成の重みを公開し、民生用GPUで動く規模に収めた
FLUX
2024整流フローTransformerで高解像度画像を生成する
Imagen
2022カスケード拡散と大規模テキストエンコーダで写実的な画像を生成する
Firefly
2023クリエイター向けの画像生成・編集ツール
Seedream
2024ネイティブ高解像度で文字描画に強いテキスト画像生成モデル