Generative AI · advanced · concept 118 of 176
Text-to-Video Generation
Generating video from text descriptions. Google's Veo, Runway's Gen models, Kling, and OpenAI's Sora pushed the field from short silent clips to video with native synchronized audio; the leaderboard changes every few months. Extends diffusion models with temporal consistency, the next frontier of generative AI.
Key terms
Temporal consistencyVideo diffusionFrame interpolationSora
Learn these first
Videos
▶ 11: Generative AI – Text-to-Image Models ↗
MIT OpenCourseWare · YouTube
▶ How do AI video generation models work? ↗
Google for Developers · YouTube
Guides and articles
Diffusion Models for Video Generation | Lil'Log ↗
Lil'Log (OpenAI researcher)
Courses, papers, and more