Shared by automation-2 using Learnlo
Create your own pack βPick a topic to learn or start your exam journey.
0/20 topics mastered
Stable Diffusion is a deep learning, text-to-image generative model released in 2022 that uses diffusion techniques to create detailed images from text prompts. It is a latent diffusion model, meaning it operates in a compressed latent space using a variational autoencoder (VAE), a denoising network (U-Net or related backbone), and an optional text encoder for conditioning. The modelβs code and weights were made publicly available, enabling it to run locally on consumer hardware with modest GPU memory (as low as a few GB VRAM), which helped it spread beyond earlier proprietary cloud-only systems. Its release history includes multiple major model generations and updates. The initial release date is August 22, 2022, and the latest release mentioned in the provided content is Stable Diffusion 3.5 (model), released on October 22, 2024. The series also includes notable architectural shifts such as SD XL (with higher-resolution capability and an accompanying refiner) and SD 3.0, which changes the backbone from a U-Net-based design to a rectified-flow transformer approach. Overall, the progression reflects improvements in resolution, conditioning, and generation quality over time.
0/2 modes complete
0/2 modes complete