本文へスキップ
AI図鑑

Veo

1080pでショットの一貫した動画クリップを生成する

Google DeepMind モデル クローズド
入力テキスト画像動画

本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。

これは何か

Veo is a text-to-video model Google DeepMind announced in May 2024. It accepts text or image input, generates 1080p video longer than a minute, and supports several cinematic styles and camera controls. Veo is also used in product settings such as YouTube’s short-form tools. Its weights are not open.

なぜ覚えておく価値があるか

It put high resolution, longer duration and controllable camera work into a single video model and fed it directly into Google’s product line, making it a major entry in the 2024 text-to-video race.

主な仕様

Resolution
1080p
Duration
Over 60 seconds (first disclosure)
Input
Text, image
Open weights
No

対応する能力

関連する概念

同種の製品