OpenAI GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models @TheAIEpiphany
OpenAI GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models  @TheAIEpiphany
Uploaded December 2021 | Updated September 2026, 1 week ago
❀️ Become The AI Epiphany Patreon ❀️
patreon.com/theaiepiphany

πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦ Join our Discord community πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦
discord.gg/peBrCpheKE

In this video I cover a new paper from OpenAI - "GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models" where they combine diffusion models with transformers to outperform their older DALL-E model.

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬
βœ… GLIDE paper: arxiv.org/abs/2112.10741
βœ… GLIDE code: github.com/openai/glide-text2im

Learning about diffusion models
Papers:
βœ… Seminal (2015): arxiv.org/pdf/1503.03585.pdf
βœ… DDPM (2020): arxiv.org/pdf/2006.11239.pdf
βœ… OpenAI (1): arxiv.org/pdf/2102.09672.pdf
βœ… OpenAI (2): arxiv.org/pdf/2105.05233.pdf

Blogs:
βœ… Score-based models: yang-song.github.io/blog/2021/score
βœ… Diffusion models: lilianweng.github.io/lil-log/2021/07/11/diffusion-models.html
β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

⌚️ Timetable:
00:00 Intro to GLIDE - results
04:00 Intro to diffusion models
07:10 Inpainting and other awesome results
11:05 Diffusion models in depth
20:45 VAE inspired loss
31:30 GLIDE pipeline (diffusion + transformers)
34:15 Guided diffusion
38:00 Classifier-free guidance
42:25 CLIP guidance
45:25 Comparison with other models
48:30 Safety considerations
49:25 Failure cases
51:40 Outro

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬
πŸ’° BECOME A PATREON OF THE AI EPIPHANY ❀️

If these videos, GitHub projects, and blogs help you,
consider helping me out by supporting me on Patreon!

The AI Epiphany - patreon.com/theaiepiphany
One-time donation - paypal.com/paypalme/theaiepiphany

Huge thank you to these AI Epiphany patreons:
Eli Mahler
Kulsoom Abdullah
Petar VeličkoviΔ‡

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

πŸ’Ό LinkedIn - linkedin.com/in/aleksagordic
🐦 Twitter - twitter.com/gordic_aleksa
πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦ Discord - discord.gg/peBrCpheKE

πŸ“Ί YouTube - youtube.com/c/TheAIEpiphany
πŸ“š Medium - gordicaleksa.medium.com
πŸ’» GitHub - github.com/gordicaleksa
πŸ“’ AI Newsletter - aiepiphany.substack.com

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

#glide #openai #diffusionmodels
OpenAI GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion ModelsHigh Fidelity Neural Audio Compression | Paper & Code ExplainedDay 24: Open NLLB - back from China, filtering HBS data (Pt 3)When Vision Transformers Outperform ResNets without Pretraining | Paper ExplainedDeepMind DetCon: Efficient Visual Pretraining with Contrastive Detection | Paper ExplainedArchie: an engineering AGI for Dyson Spheres | P-1 AI | $23 million seed roundDay 13: Open NLLB - Aya, dedup sharding analysis, analyzing training (Pt 2.)BigScience BLOOM | 3D Parallelism Explained | Large Language Models | ML Coding SeriesOpenAI DALL-E 3 with James Betker (1st author)Day 10: Open NLLB - evaluation data, filtering (Pt 3.)Day 13: Open NLLB - FSDP paper, kicking off the 1st run (Pt 1.)Day 8: Meta NLLB - analyzing the training script, Jais paper (Pt. 1)
Aleksa Gordić - The AI Epiphany |

OpenAI GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER