입력텍스트이미지영상
이 페이지의 본문은 영어로 제공됩니다. 제목과 요약은 한국어로 번역되었습니다.
무엇인가
Veo is a text-to-video model Google DeepMind announced in May 2024. It accepts text or image input, generates 1080p video longer than a minute, and supports several cinematic styles and camera controls. Veo is also used in product settings such as YouTube’s short-form tools. Its weights are not open.
기억할 만한 이유
It put high resolution, longer duration and controllable camera work into a single video model and fed it directly into Google’s product line, making it a major entry in the 2024 text-to-video race.
주요 사양
- Resolution
- 1080p
- Duration
- Over 60 seconds (first disclosure)
- Input
- Text, image
- Open weights
- No
소속 능력
관련 개념
동종 제품
Sora
2024OpenAI
한 줄 설명으로 최대 약 1분 길이의 일관된 영상을 만든다
모델 클로즈드 소스
텍스트이미지영상
Gen-3 Alpha
2024Runway
영상·광고를 겨냥한 제어성 높은 텍스트-비디오 모델
모델 클로즈드 소스
텍스트이미지영상
Kling
2024Kuaishou (Kling)
텍스트와 이미지 모두로 영상을 만드는 숏폼 모델
모델 클로즈드 소스
텍스트이미지영상
Dream Machine
2024Luma AI
텍스트나 이미지로 움직임이 있는 짧은 영상을 생성한다
모델 클로즈드 소스
텍스트이미지영상
Seedance
2024ByteDance (Seed)
멀티 숏 내러티브를 겨냥한 영상 생성 모델
모델 클로즈드 소스
텍스트이미지영상
Synthesia
2019Synthesia
글자를 입력하면 아바타가 말하는 영상이 나온다
앱 클로즈드 소스
텍스트영상