Computer Vision · intermediate · concept 103 of 176
Image Generation
Creating new images from text descriptions or other inputs. Stable Diffusion, Midjourney, FLUX, and the image models built into ChatGPT and Gemini can generate photorealistic images from text prompts.
Key terms
Text-to-imageDiffusionCLIP guidanceInpainting
Learn these first
Videos
▶ Large Language Models explained briefly ↗
3Blue1Brown · YouTube
▶ Stanford CS25: V5 I Transformers in Diffusion Models for Image Generation and Beyond ↗
Stanford Online · YouTube
Guides and articles
What are Diffusion Models? | Lil'Log ↗
Lil'Log (OpenAI researcher)
Diffusers · Hugging Face ↗
Hugging Face
Courses, papers, and more
This unlocks