入力テキスト画像動画
本ページの本文は英語で提供されています。タイトルと導入は日本語化されています。
これは何か
Kling is a video-generation model Kuaishou released in June 2024. It supports both text-to-video and image-to-video, produces clips of 5 or 10 seconds up to 1080p, and emphasises motion magnitude and visual steadiness. It was opened to users through a web product and does not disclose its weights.
なぜ覚えておく価値があるか
It was an early product to offer both text-to-video and image-to-video to the public and gained wide use in 2024 for its larger motion range and steadiness.
主な仕様
- Resolution
- Up to 1080p
- Duration
- 5 or 10 seconds
- Input
- Text, image
- Open weights
- No
対応する能力
関連する概念
同種の製品
Sora
2024OpenAI
一文の説明から最長約1分の一貫した動画を生成する
モデル クローズド
テキスト画像動画
Veo
2024Google DeepMind
1080pでショットの一貫した動画クリップを生成する
モデル クローズド
テキスト画像動画
Gen-3 Alpha
2024Runway
映像・広告向けの制御性の高いテキスト動画生成モデル
モデル クローズド
テキスト画像動画
Hailuo
2024MiniMax
指示追従とカメラワークを重視したショート動画モデル
モデル クローズド
テキスト画像動画
Stable Video Diffusion
2023Stability AI
拡散モデルで一枚の静止画を短い動画に変える
モデル オープンウェイト
画像動画
Dream Machine
2024Luma AI
テキストや画像から動きのある短い動画を生成する
モデル クローズド
テキスト画像動画
Seedance
2024ByteDance (Seed)
マルチショットの物語表現を狙った動画生成モデル
モデル クローズド
テキスト画像動画
SenseAvatar
2022SenseTime
一枚の人物画像と音声からリップシンクしたデジタルヒューマン動画を生成する
API クローズド
画像音声動画