본문으로 건너뛰기
AI 도감

DALL·E 3

긴 프롬프트를 상세한 설명으로 다시 쓴 뒤 이미지를 생성한다

OpenAI 모델 클로즈드 소스
입력텍스트이미지

이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.

무엇인가

DALL·E 3 is OpenAI’s text-to-image model, released in October 2023. It folds prompt rewriting and image generation into one flow: the model first expands a short user prompt into a more detailed scene description, then generates the image from it, which improves adherence to long prompts and complex constraints. Training used automatic recaptioning to add fine-grained descriptions to images, narrowing the gap between images and their text labels. It is available through the API and ChatGPT.

기억할 만한 이유

It made “understand the prompt first, then draw” a single pipeline, markedly improving adherence to long prompts and constraints, and turned conversational image generation into a routine ChatGPT capability.

주요 사양

Resolution
1024×1024, 1792×1024, 1024×1792
Prompt handling
Automatic recaptioning during training
Open weights
No
Availability
API and ChatGPT

소속 능력

관련 개념

동종 제품