이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.
무엇인가
DALL·E 3 is OpenAI’s text-to-image model, released in October 2023. It folds prompt rewriting and image generation into one flow: the model first expands a short user prompt into a more detailed scene description, then generates the image from it, which improves adherence to long prompts and complex constraints. Training used automatic recaptioning to add fine-grained descriptions to images, narrowing the gap between images and their text labels. It is available through the API and ChatGPT.
기억할 만한 이유
It made “understand the prompt first, then draw” a single pipeline, markedly improving adherence to long prompts and constraints, and turned conversational image generation into a routine ChatGPT capability.
주요 사양
- Resolution
- 1024×1024, 1792×1024, 1024×1792
- Prompt handling
- Automatic recaptioning during training
- Open weights
- No
- Availability
- API and ChatGPT
소속 능력
관련 개념
동종 제품
Midjourney
2022미적 스타일로 알려진 텍스트-이미지 서비스
Stable Diffusion
2022텍스트-이미지 가중치를 공개하고 소비자용 GPU에서 돌아가게 만들었다
FLUX
2024정류 흐름 트랜스포머로 고해상도 이미지를 생성한다
Imagen
2022계단식 확산과 대형 텍스트 인코더로 사실적인 이미지를 생성한다
Firefly
2023크리에이터를 위한 이미지 생성·편집 도구
Seedream
2024네이티브 고해상도와 뛰어난 문자 렌더링의 텍스트-이미지 모델