本文へスキップ
AI図鑑

Stable Video Diffusion

拡散モデルで一枚の静止画を短い動画に変える

Stability AI モデル オープンウェイト
入力画像動画

本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。

これは何か

Stable Video Diffusion (SVD) is an image-to-video model Stability AI released in November 2023 with open weights. It frames image-to-video generation as latent-space diffusion: conditioned on a single still image, it produces a short clip with an adjustable amount of motion. The model generates 14 frames at 576×1024 resolution by default. It shipped as both an image-to-video model and a multi-view model, the latter synthesising views orbiting an object from a single image.

なぜ覚えておく価値があるか

It was one of the few image-to-video models to release open weights, letting “make a short video from one image” run locally or on self-hosted services and advancing open practice in image-to-video.

主な仕様

Resolution
576×1024
Default frames
14 frames
Input
Single image
Open weights
Yes

対応する能力

関連する概念

同種の製品