WHAT IT IS
Gen-3 Alpha is a text-to-video model Runway released in June 2024. It accepts text and image input, generates clips of about 5 to 10 seconds at 720p, and offers controls for camera movement and style. Runway positions it for professional settings such as film and advertising, used alongside its video-editing tools. The model does not release its weights.
Why it matters
It positioned video generation squarely inside professional creative workflows, emphasising controllability and integration with editing tools, as text-to-video moved toward commercial use in 2024.
Key specs
- Resolution
- 720p
- Duration
- About 5 to 10 seconds
- Input
- Text, image
- Open weights
- No
Capabilities
Related concepts
Diffusion Models
Learn a thousand tiny denoising steps, and you can build an image from pure noise
Multimodal Generation
One model that learns to speak, to draw, to move — even to model the 3D world
Training & Inference Infrastructure
Memory decides how large a model you can train, communication how long it takes — raw compute is rarely the bottleneck
Comparable products
Sora
2024Generates coherent video up to about a minute long from a description
Kling
2024A short-video model for both text-to-video and image-to-video
Dream Machine
2024Generates short videos with motion from text or an image
Veo
2024Generates 1080p video clips with coherent shots
Stable Video Diffusion
2023Turns a single still image into a short video with a diffusion model
Hailuo
2024A short-video model focused on instruction following and camera language