The Most Absurd Way To Train LLMs... With 3x Less Memory!? @bycloudAI
The Most Absurd Way To Train LLMs... With 3x Less Memory!?  @bycloudAI
Uploaded August 2026 | Updated September 2026, 1 week ago
Need high quality cloud GPUs? Check out Verda now, and use the code BYCLOUD-50 to get $50 of compute credits for just $5!
verda.com/?utm_source=bycloud&utm_medium=referral&utm_campaign=sponsorship&utm_content=diffusion_blocks

A great news for us gpu poors? This research DiffusionBlocks shows that you can train LLMs with 2-3x less memory with minimal performance loss. If crazier, you could even go up to 4x or 6x!


my latest project: Intuitive AI Academy
We just wrote a new piece on Optimization!!
https://intuitiveai.academy/
limited time code "LOCKIN" for 35% off yearly plan

My Newsletter (weekly top research papers)
mail.bycloud.ai

My Patreon
patreon.com/c/bycloud


DiffusionBlocks
[Paper] arxiv.org/abs/2506.14202
[Project Page] pub.sakana.ai/diffusionblocks
[Code] github.com/SakanaAI/DiffusionBlocks



Try out my new fav place to learn how to code scrimba.com/?via=bycloudAI

This video is supported by the kind Patrons & YouTube Members:
🙏Spam Maj, Alex, Chris LeDoux, DX Research Group, Poof N' Inu, Deagan, Robert Zawiasa, Ryszard Warzocha, Tobe2d, Louis Muk, Akkusativ, Kevin Tai, Mark Buckler, NO U, Tony Jimenez, Ângelo Fonseca, jiye, Anushka, Asad Dhamani, Binnie Yiu, Calvin Yan, Clayton Ford, Diego Silva, Etrotta, Gonzalo Fidalgo, Handenon, Hector, Jake Disco very, Michael Brenner, Nilly K, OlegWock, Daddy Wen, Shuhong Chen, Sid_Cipher, Stefan Lorenz, Sup, tantan assawade, Thipok Tham, Thomas Di Martino, Thomas Lin, Richárd Nagyfi, Paperboy, mika, Leo, Berhane-Meskel, Kadhai Pesalam, mayssam, Bill Mangrum, nyaa, Toru Mon, Lame Plane, Matej Macak, Len Mo, saylikhapekar, ZyanSheep, THEVIERAOS, Ricardo Raphael Corona-Moreno, superchordate


[Discord] discord.gg/NhJZGtH
[Twitter] twitter.com/bycloudai
[Patreon] patreon.com/bycloud
[Business Inquiries] bycloud@smoothmedia.co
[Other Inquiries] bycloudai@gmail.com
[Profile & Banner Art] twitter.com/pygm7
[Video Editor] @aduckchicken2
Manim Animations created with Manimate manimate.ai
[Ko-fi] ko-fi.com/bycloudai
The Most Absurd Way To Train LLMs... With 3x Less Memory!?How Googles Transformer 2.0 Might Be The AI Breakthrough We NeedLLMs organizes knowledge into shapes...? #llm #ai #airesearchThe Chinese AI IcebergClaude 3.7 Sonnet: The AI Code King Has Returned (within 1 week)The LK-99 of AI: The Reflection-70B Controversy Full RundownAll You Need To Know About Running LLMs LocallyThe Difference of AI Videos No One Tells You AboutAttention Sink: The Fluke That Made LLMs Actually UsableThe AI Hardware Arms Race Is Getting Out of HandGoogles New OS Gemma 4 Series Beats Models 10x Its Size!Has ChatGPT Finally Been Dethroned? Claude 3 Review
bycloud |

The Most Absurd Way To Train LLMs... With 3x Less Memory!?

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER