OpenAI GLIDE (Diffusion) | ML Coding series | Towards Photorealistic Image Generation and Editing @TheAIEpiphany
OpenAI GLIDE (Diffusion) | ML Coding series | Towards Photorealistic Image Generation and Editing  @TheAIEpiphany
Uploaded July 2022 | Updated September 2026, 1 week ago
❀️ Become The AI Epiphany Patreon ❀️
patreon.com/theaiepiphany

πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦ Join our Discord community πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦
discord.gg/peBrCpheKE

5th video of the ML coding series! In this one I cover OpenAI's GLIDE model from the "GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models" paper.

It directly builds upon the "Diffusion Models Beat GANs on Image Synthesis" paper that I've covered in the previous video of the series. It is also a direct precursor to DALL-E 2!

I explain classifier-free guidance as well as CLIP guidance in this one.

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬
βœ… Paper: arxiv.org/abs/2112.10741
βœ… GitHub: github.com/openai/glide-text2im
β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

⌚️ Timetable:
00:00:00 [paper] GLIDE - quick overview
00:03:30 Classifier free guidance explained
00:07:25 CLIP guidance explained
00:10:49 Training details, etc.
00:14:20 [coding] CLIP guidance script
00:26:25 Generating CLIP text target embedding
00:32:32 Reverse process (p_sample_loop)
00:34:35 How is text conditioning implemented
00:43:00 CLIP guidance code
00:46:26 Classifier-free guidance script
00:51:47 Main logic
00:56:35 Showing generated images
00:58:23 Outro

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬
πŸ’° BECOME A PATREON OF THE AI EPIPHANY ❀️

If these videos, GitHub projects, and blogs help you,
consider helping me out by supporting me on Patreon!

The AI Epiphany - patreon.com/theaiepiphany
One-time donation - paypal.com/paypalme/theaiepiphany

Huge thank you to these AI Epiphany patreons:
Eli Mahler
Petar VeličkoviΔ‡

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

πŸ’Ό LinkedIn - linkedin.com/in/aleksagordic
🐦 Twitter - twitter.com/gordic_aleksa
πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦ Discord - discord.gg/peBrCpheKE

πŸ“Ί YouTube - youtube.com/c/TheAIEpiphany
πŸ“š Medium - gordicaleksa.medium.com
πŸ’» GitHub - github.com/gordicaleksa
πŸ“’ AI Newsletter - aiepiphany.substack.com

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

#glide #diffusion #imagesynthesis
OpenAI GLIDE (Diffusion) | ML Coding series | Towards Photorealistic Image Generation and EditingDay 29: Open NLLB - training fasttext LID (Pt 1)Fine-tune LLMs 30x faster! With Daniel Han (Unsloth AI)Day 3: Replicating Metas NLLB - primary & mined data (Pt. 2)Ishan Misra (Meta) - Emu Video GenerationAndrew Huberman transcripts | MLOps series announcementDay 18: Open NLLB - Serbian parallel corpora, paper reading, batch analysis (Pt 2)Day 28: Open NLLB - debugging fuzzy dedup, training fasttext LID (Pt 3)Lucas Beyer (Google DeepMind) - Convergence of Vision & LanguageGet Started With Stable Diffusion! (Code, HF Spaces, Diffusers Notebooks)Stable Diffusion: High-Resolution Image Synthesis with Latent Diffusion Models | ML Coding SeriesDay 12: Open NLLB - on-boarding doc for new-joiners, eval (Pt 2.)
Aleksa Gordić - The AI Epiphany |

OpenAI GLIDE (Diffusion) | ML Coding series | Towards Photorealistic Image Generation and Editing

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER