मुख्य सामग्री पर जाएँ

Sora

एक विवरण से लगभग एक मिनट तक का सुसंगत वीडियो बनाता है

OpenAI मॉडल बंद स्रोत
इनपुटटेक्स्टइमेजवीडियो

यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।

यह क्या है

Sora is a text-to-video model OpenAI first unveiled in February 2024. It tokenises video into spacetime patches and performs diffusion-style generation over that unified representation, which lets it accept different resolutions and durations. It accepts both text and image input for text-to-video and image-to-video. In its first demonstrations, clips ran up to about 60 seconds at 1080p, making it one of the few models to show minute-scale coherent video at the time.

यह क्यों महत्वपूर्ण है

It brought minute-scale, shot-coherent video generation into public view for the first time, moving video generation beyond short clips toward consistency over longer stretches.

मुख्य विशिष्टताएँ

Resolution
Up to 1080p (first disclosure)
Duration
Up to about 60 seconds (first disclosure)
Input
Text, image
Open weights
No

संबंधित क्षमताएँ

संबंधित अवधारणाएँ

समान उत्पाद

Veo

2024
Google DeepMind

सुसंगत शॉट के साथ 1080p वीडियो क्लिप बनाता है

मॉडल बंद स्रोत
टेक्स्टइमेजवीडियो

Gen-3 Alpha

2024
Runway

फ़िल्म और विज्ञापन के लिए अत्यधिक नियंत्रण योग्य टेक्स्ट-टू-वीडियो मॉडल

मॉडल बंद स्रोत
टेक्स्टइमेजवीडियो

Kling

2024
Kuaishou (Kling)

टेक्स्ट-टू-वीडियो और इमेज-टू-वीडियो दोनों के लिए शॉर्ट-वीडियो मॉडल

मॉडल बंद स्रोत
टेक्स्टइमेजवीडियो

Hailuo

2024
MiniMax

निर्देश-पालन और कैमरा भाषा पर केंद्रित शॉर्ट-वीडियो मॉडल

मॉडल बंद स्रोत
टेक्स्टइमेजवीडियो

Stable Video Diffusion

2023
Stability AI

विसरण मॉडल से एक स्थिर चित्र को छोटे वीडियो में बदलता है

मॉडल खुले वेट
इमेजवीडियो

Dream Machine

2024
Luma AI

पाठ या चित्र से गति युक्त छोटे वीडियो बनाता है

मॉडल बंद स्रोत
टेक्स्टइमेजवीडियो

Seedance

2024
ByteDance (Seed)

मल्टी-शॉट कथा-कथन के लिए वीडियो जनरेटिंग मॉडल

मॉडल बंद स्रोत
टेक्स्टइमेजवीडियो

Synthesia

2019
Synthesia

टेक्स्ट लिखें, बोलते अवतार वाला वीडियो पाएँ

ऐप बंद स्रोत
टेक्स्टवीडियो