Ir para o conteúdo
Atlas de IA

Sora

Gera vídeo coerente de até cerca de um minuto a partir de uma descrição

OpenAI Modelo Fechado
entradaTextoImagemVídeo

O texto completo é apresentado em inglês; o título e o resumo estão traduzidos.

O QUE É

Sora is a text-to-video model OpenAI first unveiled in February 2024. It tokenises video into spacetime patches and performs diffusion-style generation over that unified representation, which lets it accept different resolutions and durations. It accepts both text and image input for text-to-video and image-to-video. In its first demonstrations, clips ran up to about 60 seconds at 1080p, making it one of the few models to show minute-scale coherent video at the time.

Por que vale a pena lembrar

It brought minute-scale, shot-coherent video generation into public view for the first time, moving video generation beyond short clips toward consistency over longer stretches.

Especificações-chave

Resolution
Up to 1080p (first disclosure)
Duration
Up to about 60 seconds (first disclosure)
Input
Text, image
Open weights
No

Capacidades

Conceitos relacionados

Produtos semelhantes