本文へスキップ
AI図鑑

Sora

一文の説明から最長約1分の一貫した動画を生成する

OpenAI モデル クローズド
入力テキスト画像動画

本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。

これは何か

Sora is a text-to-video model OpenAI first unveiled in February 2024. It tokenises video into spacetime patches and performs diffusion-style generation over that unified representation, which lets it accept different resolutions and durations. It accepts both text and image input for text-to-video and image-to-video. In its first demonstrations, clips ran up to about 60 seconds at 1080p, making it one of the few models to show minute-scale coherent video at the time.

なぜ覚えておく価値があるか

It brought minute-scale, shot-coherent video generation into public view for the first time, moving video generation beyond short clips toward consistency over longer stretches.

主な仕様

Resolution
Up to 1080p (first disclosure)
Duration
Up to about 60 seconds (first disclosure)
Input
Text, image
Open weights
No

対応する能力

関連する概念

同種の製品