本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。
これは何か
Sora is a text-to-video model OpenAI first unveiled in February 2024. It tokenises video into spacetime patches and performs diffusion-style generation over that unified representation, which lets it accept different resolutions and durations. It accepts both text and image input for text-to-video and image-to-video. In its first demonstrations, clips ran up to about 60 seconds at 1080p, making it one of the few models to show minute-scale coherent video at the time.
なぜ覚えておく価値があるか
It brought minute-scale, shot-coherent video generation into public view for the first time, moving video generation beyond short clips toward consistency over longer stretches.
主な仕様
- Resolution
- Up to 1080p (first disclosure)
- Duration
- Up to about 60 seconds (first disclosure)
- Input
- Text, image
- Open weights
- No
対応する能力
関連する概念
同種の製品
Veo
20241080pでショットの一貫した動画クリップを生成する
Gen-3 Alpha
2024映像・広告向けの制御性の高いテキスト動画生成モデル
Kling
2024テキストからも画像からも動画を生成できるショート動画モデル
Hailuo
2024指示追従とカメラワークを重視したショート動画モデル
Stable Video Diffusion
2023拡散モデルで一枚の静止画を短い動画に変える
Dream Machine
2024テキストや画像から動きのある短い動画を生成する
Seedance
2024マルチショットの物語表現を狙った動画生成モデル
Synthesia
2019テキストを入力するとアバターが話す動画ができる