本文へスキップ
AI図鑑

DALL·E 3

長い指示を詳細な説明に書き換えてから画像を生成する

OpenAI モデル クローズド
入力テキスト画像

本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。

これは何か

DALL·E 3 is OpenAI’s text-to-image model, released in October 2023. It folds prompt rewriting and image generation into one flow: the model first expands a short user prompt into a more detailed scene description, then generates the image from it, which improves adherence to long prompts and complex constraints. Training used automatic recaptioning to add fine-grained descriptions to images, narrowing the gap between images and their text labels. It is available through the API and ChatGPT.

なぜ覚えておく価値があるか

It made “understand the prompt first, then draw” a single pipeline, markedly improving adherence to long prompts and constraints, and turned conversational image generation into a routine ChatGPT capability.

主な仕様

Resolution
1024×1024, 1792×1024, 1024×1792
Prompt handling
Automatic recaptioning during training
Open weights
No
Availability
API and ChatGPT

対応する能力

関連する概念

同種の製品