Chuyển đến nội dung
Bản đồ AI

Stable Diffusion

Mở công khai trọng số tạo ảnh từ văn bản, đủ nhẹ để chạy trên GPU phổ thông

Stability AI Mô hình Trọng số mở
đầu vàoVăn bảnHình ảnhHình ảnh

Toàn văn của trang này được trình bày bằng tiếng Anh; tiêu đề và phần dẫn đã được bản địa hóa.

NÓ LÀ GÌ

Stable Diffusion is an open-source text-to-image model that Stability AI released in August 2022. It runs the diffusion process in the latent space of a pretrained autoencoder, sharply cutting computation so the model can run on consumer GPUs. Text is turned into conditioning vectors by a CLIP text encoder and injected into the denoising network through cross-attention. The first 1.x version natively outputs 512×512, with a U-Net of roughly 860 million parameters and weights released under an open licence.

Vì sao đáng ghi nhớ

By open-sourcing text-to-image weights and making them runnable on consumer GPUs, it directly spawned the whole open-source image ecosystem of fine-tunes, plugins and local generation, marking the point where generative imagery reached a broad audience.

Thông số chính

Native resolution
512×512 (1.x)
U-Net parameters
About 860M
Architecture
Latent diffusion + CLIP text encoder
Open weights
Yes
Released
2022-08

Năng lực liên quan

Khái niệm liên quan

Sản phẩm cùng loại