تخطٍّ إلى المحتوى
أطلس الذكاء الاصطناعي

Veo

يولّد مقاطع فيديو بدقة 1080p بمشاهد متماسكة

Google DeepMind نموذج مغلق
إدخالنصصورةفيديو

يُعرض النص الكامل باللغة الإنجليزية؛ وقد تمت ترجمة العنوان والملخص.

ما هو

Veo is a text-to-video model Google DeepMind announced in May 2024. It accepts text or image input, generates 1080p video longer than a minute, and supports several cinematic styles and camera controls. Veo is also used in product settings such as YouTube’s short-form tools. Its weights are not open.

لماذا يستحق التذكّر

It put high resolution, longer duration and controllable camera work into a single video model and fed it directly into Google’s product line, making it a major entry in the 2024 text-to-video race.

المواصفات الأساسية

Resolution
1080p
Duration
Over 60 seconds (first disclosure)
Input
Text, image
Open weights
No

القدرات المرتبطة

المفاهيم ذات الصلة

منتجات منافسة