본문으로 건너뛰기
AI 도감

Veo

1080p의 일관된 숏 영상을 생성한다

Google DeepMind 모델 클로즈드 소스
입력텍스트이미지영상

이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.

무엇인가

Veo is a text-to-video model Google DeepMind announced in May 2024. It accepts text or image input, generates 1080p video longer than a minute, and supports several cinematic styles and camera controls. Veo is also used in product settings such as YouTube’s short-form tools. Its weights are not open.

기억할 만한 이유

It put high resolution, longer duration and controllable camera work into a single video model and fed it directly into Google’s product line, making it a major entry in the 2024 text-to-video race.

주요 사양

Resolution
1080p
Duration
Over 60 seconds (first disclosure)
Input
Text, image
Open weights
No

소속 능력

관련 개념

동종 제품