본문으로 건너뛰기
AI 도감

Sora

한 줄 설명으로 최대 약 1분 길이의 일관된 영상을 만든다

OpenAI 모델 클로즈드 소스
입력텍스트이미지영상

이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.

무엇인가

Sora is a text-to-video model OpenAI first unveiled in February 2024. It tokenises video into spacetime patches and performs diffusion-style generation over that unified representation, which lets it accept different resolutions and durations. It accepts both text and image input for text-to-video and image-to-video. In its first demonstrations, clips ran up to about 60 seconds at 1080p, making it one of the few models to show minute-scale coherent video at the time.

기억할 만한 이유

It brought minute-scale, shot-coherent video generation into public view for the first time, moving video generation beyond short clips toward consistency over longer stretches.

주요 사양

Resolution
Up to 1080p (first disclosure)
Duration
Up to about 60 seconds (first disclosure)
Input
Text, image
Open weights
No

소속 능력

관련 개념

동종 제품