ConvNeXt: A ConvNet for the 2020s | Paper Explained @TheAIEpiphany
ConvNeXt: A ConvNet for the 2020s | Paper Explained  @TheAIEpiphany
Uploaded January 2022 | Updated September 2026, 1 week ago
❀️ Become The AI Epiphany Patreon ❀️
patreon.com/theaiepiphany

πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦ Join our Discord community πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦
discord.gg/peBrCpheKE

In this video I cover the recently published "A ConvNet for the 2020s" paper. They show that ConvNets are still in the game! - by adding new design ideas and training procedures they outperform vision transformers even in big data regimes and without any attention layers.

Convolutional prior continues to stand the test of time in the field of computer vision.

Note: I also partially cover the Swin transformer paper in case you missed out on that one. :)

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬
βœ… Paper: arxiv.org/abs/2201.03545
βœ… GitHub: github.com/facebookresearch/ConvNeXt
β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

⌚️ Timetable:
00:00 Intro - convergence of transformers and CNNs
05:05 Main diagram explained
07:40 Main diagram corrections
10:10 Swin transformer recap
20:20 Modernizing ResNets
24:10 Diving deeper: stage ratio
27:20 Diving deeper: misc (inverted bottleneck, depthwise conv...)
34:45 Results (classification, object detection, segmentation)
37:35 RIP DanNet
38:40 Summary and outro

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬
πŸ’° BECOME A PATREON OF THE AI EPIPHANY ❀️

If these videos, GitHub projects, and blogs help you,
consider helping me out by supporting me on Patreon!

The AI Epiphany - patreon.com/theaiepiphany
One-time donation - paypal.com/paypalme/theaiepiphany

Huge thank you to these AI Epiphany patreons:
Eli Mahler
Petar VeličkoviΔ‡

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

πŸ’Ό LinkedIn - linkedin.com/in/aleksagordic
🐦 Twitter - twitter.com/gordic_aleksa
πŸ‘¨β€πŸ‘©β€πŸ‘§β€πŸ‘¦ Discord - discord.gg/peBrCpheKE

πŸ“Ί YouTube - youtube.com/c/TheAIEpiphany
πŸ“š Medium - gordicaleksa.medium.com
πŸ’» GitHub - github.com/gordicaleksa
πŸ“’ AI Newsletter - aiepiphany.substack.com

β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬β–¬

#convnext #visiontransformers #computervision
ConvNeXt: A ConvNet for the 2020s | Paper ExplainedVQ-GAN: Taming Transformers for High-Resolution Image Synthesis | Paper ExplainedDay 14: Open NLLB - exploring BLEU, chrF++, logging (Pt 3. cont.)DALL-E: Zero-Shot Text-to-Image Generation | Paper ExplainedDay 24: Open NLLB - back from China, analyzing spikes, preparing HBS run (Pt 2)OpenAI CLIP | Machine Learning Coding SeriesFake It Till You Make It (Microsoft) | Paper ExplainedLLaMA 2 w/ Thomas Scialom (LLaMA 2 lead)Day 24: Open NLLB - back from China, fuzzy dedup, preparing HBS run (Pt 1)Day 12: Open NLLB - on-boarding doc for new-joiners, eval (Pt 2.)Day 22: Open NLLB - HBS data analysis, split into Cyrillic & Latin (Pt 2)Day 4: Training 600M NLLB - data preps (Pt. 2)
Aleksa Gordić - The AI Epiphany |

ConvNeXt: A ConvNet for the 2020s | Paper Explained

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER