यह पृष्ठ अंग्रेज़ी में प्रस्तुत है; शीर्षक और सारांश का स्थानीयकरण किया गया है।
यह क्या है
Veo is a text-to-video model Google DeepMind announced in May 2024. It accepts text or image input, generates 1080p video longer than a minute, and supports several cinematic styles and camera controls. Veo is also used in product settings such as YouTube’s short-form tools. Its weights are not open.
यह क्यों महत्वपूर्ण है
It put high resolution, longer duration and controllable camera work into a single video model and fed it directly into Google’s product line, making it a major entry in the 2024 text-to-video race.
मुख्य विशिष्टताएँ
- Resolution
- 1080p
- Duration
- Over 60 seconds (first disclosure)
- Input
- Text, image
- Open weights
- No
संबंधित क्षमताएँ
संबंधित अवधारणाएँ
विसरण मॉडल
हज़ार छोटे शोर-हटाने के चरण सीखें, और शुद्ध शोर से चित्र बना सकेंगे
बहु-मॉडल जनरेशन
एक ही मॉडल बोलना, चित्र बनाना, हिलना, और यहाँ तक कि 3D संसार का नमूना बनाना सीखता है
प्रशिक्षण एवं अनुमान अधोसंरचना
स्मृति तय करती है कि कितना बड़ा मॉडल प्रशिक्षित हो सकता है, और संचार तय करता है कि कितना समय लगेगा
समान उत्पाद
Sora
2024एक विवरण से लगभग एक मिनट तक का सुसंगत वीडियो बनाता है
Gen-3 Alpha
2024फ़िल्म और विज्ञापन के लिए अत्यधिक नियंत्रण योग्य टेक्स्ट-टू-वीडियो मॉडल
Kling
2024टेक्स्ट-टू-वीडियो और इमेज-टू-वीडियो दोनों के लिए शॉर्ट-वीडियो मॉडल
Dream Machine
2024पाठ या चित्र से गति युक्त छोटे वीडियो बनाता है
Seedance
2024मल्टी-शॉट कथा-कथन के लिए वीडियो जनरेटिंग मॉडल
Synthesia
2019टेक्स्ट लिखें, बोलते अवतार वाला वीडियो पाएँ