Ir para o conteúdo
Atlas de IA

DALL·E 3

Reescreve um prompt longo em descrição detalhada e depois desenha a imagem

OpenAI Modelo Fechado
entradaTextoImagem

O texto completo é apresentado em inglês; o título e o resumo estão traduzidos.

O QUE É

DALL·E 3 is OpenAI’s text-to-image model, released in October 2023. It folds prompt rewriting and image generation into one flow: the model first expands a short user prompt into a more detailed scene description, then generates the image from it, which improves adherence to long prompts and complex constraints. Training used automatic recaptioning to add fine-grained descriptions to images, narrowing the gap between images and their text labels. It is available through the API and ChatGPT.

Por que vale a pena lembrar

It made “understand the prompt first, then draw” a single pipeline, markedly improving adherence to long prompts and constraints, and turned conversational image generation into a routine ChatGPT capability.

Especificações-chave

Resolution
1024×1024, 1792×1024, 1024×1792
Prompt handling
Automatic recaptioning during training
Open weights
No
Availability
API and ChatGPT

Capacidades

Conceitos relacionados

Produtos semelhantes