이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.
무엇인가
Sora is a text-to-video model OpenAI first unveiled in February 2024. It tokenises video into spacetime patches and performs diffusion-style generation over that unified representation, which lets it accept different resolutions and durations. It accepts both text and image input for text-to-video and image-to-video. In its first demonstrations, clips ran up to about 60 seconds at 1080p, making it one of the few models to show minute-scale coherent video at the time.
기억할 만한 이유
It brought minute-scale, shot-coherent video generation into public view for the first time, moving video generation beyond short clips toward consistency over longer stretches.
주요 사양
- Resolution
- Up to 1080p (first disclosure)
- Duration
- Up to about 60 seconds (first disclosure)
- Input
- Text, image
- Open weights
- No
소속 능력
관련 개념
동종 제품
Veo
20241080p의 일관된 숏 영상을 생성한다
Gen-3 Alpha
2024영상·광고를 겨냥한 제어성 높은 텍스트-비디오 모델
Kling
2024텍스트와 이미지 모두로 영상을 만드는 숏폼 모델
Hailuo
2024지시 준수와 카메라 워크에 집중한 숏폼 모델
Stable Video Diffusion
2023확산 모델로 한 장의 정지 이미지를 짧은 영상으로 만든다
Dream Machine
2024텍스트나 이미지로 움직임이 있는 짧은 영상을 생성한다
Seedance
2024멀티 숏 내러티브를 겨냥한 영상 생성 모델
Synthesia
2019글자를 입력하면 아바타가 말하는 영상이 나온다