
Stable Diffusion generates images from text descriptions
Stable Diffusion generates images from text descriptions
The model was developed by researchers from the CompVis Group at LMU Munich and Runway, with computational support from Stability AI. This collaboration resulted in a publicly accessible model that can run on consumer hardware with modest GPU capabilities. This accessibility marks a significant advancement over previous proprietary models like DALL-E and Midjourney.
Example
A user inputs "a sunset over the mountains" and receives an image of a beautiful sunset scene with mountains in the background.
Remember this
Stable Diffusion's ability to generate images from text descriptions can significantly enhance creative processes and applications in various fields.
Text adapted from Wikipedia, licensed under CC BY-SA 4.0.
the reverse process learns: p_θ(x_{t-1}|x_t)
Can we trace back the roots of life?
DDPM stands for: Denoising Diffusion Probabilistic Model
Can you imagine creating perfect photos from scratch?
Diffusion
Can you imagine a crowd spreading out evenly in a room?
Contrastive Language–Image Pre-training
CLIP embeds images and text into a shared space using contrastive learning
denoising score matching does: learns to denoise, which equals learning the score
Can you clean up a noisy picture?
2022 in science
Why do Transformers sometimes seem to 'ignore' irrelevant parts of the input?
Swipe through 100 ML concepts daily
Open Pocket Polymath