यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।
यह क्या है
Sora is a text-to-video model OpenAI first unveiled in February 2024. It tokenises video into spacetime patches and performs diffusion-style generation over that unified representation, which lets it accept different resolutions and durations. It accepts both text and image input for text-to-video and image-to-video. In its first demonstrations, clips ran up to about 60 seconds at 1080p, making it one of the few models to show minute-scale coherent video at the time.
यह क्यों महत्वपूर्ण है
It brought minute-scale, shot-coherent video generation into public view for the first time, moving video generation beyond short clips toward consistency over longer stretches.
मुख्य विशिष्टताएँ
- Resolution
- Up to 1080p (first disclosure)
- Duration
- Up to about 60 seconds (first disclosure)
- Input
- Text, image
- Open weights
- No
संबंधित क्षमताएँ
संबंधित अवधारणाएँ
विसरण मॉडल
हज़ार छोटे शोर-हटाने के चरण सीखें, और शुद्ध शोर से चित्र बना सकेंगे
बहु-मॉडल जनरेशन
एक ही मॉडल बोलना, चित्र बनाना, हिलना, और यहाँ तक कि 3D संसार का नमूना बनाना सीखता है
प्रशिक्षण एवं अनुमान अधोसंरचना
स्मृति तय करती है कि कितना बड़ा मॉडल प्रशिक्षित हो सकता है, और संचार तय करता है कि कितना समय लगेगा
समान उत्पाद
Veo
2024सुसंगत शॉट के साथ 1080p वीडियो क्लिप बनाता है
Gen-3 Alpha
2024फ़िल्म और विज्ञापन के लिए अत्यधिक नियंत्रण योग्य टेक्स्ट-टू-वीडियो मॉडल
Kling
2024टेक्स्ट-टू-वीडियो और इमेज-टू-वीडियो दोनों के लिए शॉर्ट-वीडियो मॉडल
Hailuo
2024निर्देश-पालन और कैमरा भाषा पर केंद्रित शॉर्ट-वीडियो मॉडल
Stable Video Diffusion
2023विसरण मॉडल से एक स्थिर चित्र को छोटे वीडियो में बदलता है
Dream Machine
2024पाठ या चित्र से गति युक्त छोटे वीडियो बनाता है
Seedance
2024मल्टी-शॉट कथा-कथन के लिए वीडियो जनरेटिंग मॉडल
Synthesia
2019टेक्स्ट लिखें, बोलते अवतार वाला वीडियो पाएँ