मुख्य सामग्री पर जाएँ

DALL·E 3

लंबे प्रॉम्प्ट को विस्तृत विवरण में बदलकर चित्र बनाता है

OpenAI मॉडल बंद स्रोत
इनपुटटेक्स्टइमेज

यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।

यह क्या है

DALL·E 3 is OpenAI’s text-to-image model, released in October 2023. It folds prompt rewriting and image generation into one flow: the model first expands a short user prompt into a more detailed scene description, then generates the image from it, which improves adherence to long prompts and complex constraints. Training used automatic recaptioning to add fine-grained descriptions to images, narrowing the gap between images and their text labels. It is available through the API and ChatGPT.

यह क्यों महत्वपूर्ण है

It made “understand the prompt first, then draw” a single pipeline, markedly improving adherence to long prompts and constraints, and turned conversational image generation into a routine ChatGPT capability.

मुख्य विशिष्टताएँ

Resolution
1024×1024, 1792×1024, 1024×1792
Prompt handling
Automatic recaptioning during training
Open weights
No
Availability
API and ChatGPT

संबंधित क्षमताएँ

संबंधित अवधारणाएँ

समान उत्पाद

Midjourney

2022
Midjourney

सौंदर्यपरक शैली के लिए जानी जाने वाली टेक्स्ट-टू-इमेज सेवा

ऐप बंद स्रोत
टेक्स्टइमेजइमेज

Stable Diffusion

2022
Stability AI

टेक्स्ट-टू-इमेज वेट खुले किए और उन्हें उपभोक्ता GPU पर चलने लायक बनाया

मॉडल खुले वेट
टेक्स्टइमेजइमेज

FLUX

2024
Black Forest Labs

रेक्टिफाइड-फ्लो ट्रांसफॉर्मर से उच्च-रिज़ॉल्यूशन चित्र बनाता है

मॉडल खुले वेट
टेक्स्टइमेजइमेज

Imagen

2022
Google DeepMind

कैस्केड विसरण और बड़े टेक्स्ट एनकोडर से यथार्थ चित्र बनाता है

मॉडल बंद स्रोत
टेक्स्टइमेज

Firefly

2023
Adobe

रचनाकारों के लिए चित्र निर्माण और संपादन उपकरण

ऐप बंद स्रोत
इमेजटेक्स्टइमेज

Seedream

2024
ByteDance (Seed)

मूल उच्च-रिज़ॉल्यूशन और मजबूत टेक्स्ट रेंडरिंग वाला टेक्स्ट-टू-इमेज मॉडल

मॉडल बंद स्रोत
टेक्स्टइमेज