Chuyển đến nội dung
Bản đồ AI

DALL·E 3

Viết lại câu lệnh dài thành mô tả chi tiết rồi vẽ hình ảnh

OpenAI Mô hình Đóng
đầu vàoVăn bảnHình ảnh

Toàn văn của trang này được trình bày bằng tiếng Anh; tiêu đề và phần dẫn đã được bản địa hóa.

NÓ LÀ GÌ

DALL·E 3 is OpenAI’s text-to-image model, released in October 2023. It folds prompt rewriting and image generation into one flow: the model first expands a short user prompt into a more detailed scene description, then generates the image from it, which improves adherence to long prompts and complex constraints. Training used automatic recaptioning to add fine-grained descriptions to images, narrowing the gap between images and their text labels. It is available through the API and ChatGPT.

Vì sao đáng ghi nhớ

It made “understand the prompt first, then draw” a single pipeline, markedly improving adherence to long prompts and constraints, and turned conversational image generation into a routine ChatGPT capability.

Thông số chính

Resolution
1024×1024, 1792×1024, 1024×1792
Prompt handling
Automatic recaptioning during training
Open weights
No
Availability
API and ChatGPT

Năng lực liên quan

Khái niệm liên quan

Sản phẩm cùng loại