Optimizing Model Training End-to-End: A Tiny MoE Case Study Lambda @aicouncilconf
Optimizing Model Training End-to-End: A Tiny MoE Case Study Lambda  @aicouncilconf
Uploaded June 2026 | Updated September 2026, 2 weeks ago
[2026 - DAY 3 - MODEL SYSTEMS] Cloud compute is expensive, and wasting runs on the guise of a "just scale will fix any problems" leaves you with less time to fix errors, and less compute to train the model you want. In this talk, I will discuss what are the easy optimizatiosn you might miss (minimizing communications, using the most effective algorithms, ensuring you're getting the most FLOPs possible) at the small scale, before ensuring that when you do scale up nothing is going to waste. In this particular talk, I'll be focusing on what worked at home, that then let me scale it further onto the cloud.

SPEAKER: Zach Mueller - Head of Developer Relations, Lambda

πŸ‘‰ Sign up for our "No BS" Newsletter to get the latest technical data & AI content: aicouncil.com/newsletter

ABOUT AI COUNCIL:
AI Council brings together the brightest minds in data to share industry knowledge, technical architectures and best practices in building cutting edge data & AI systems and tools.

FIND US:
Website: aicouncil.com
LinkedIn: linkedin.com/company/aicouncilconf
X: https://x.com/aicouncilconf
Optimizing Model Training End-to-End: A Tiny MoE Case Study LambdaAI Launchpad 2026: Golden AnalyticsChang She on Why He Walked Away from Parquet to Build LanceDBThe 2% gains that 10x your inference | Lessons from AWS on optimizing VLMsQ&A with Scott Breitenother, Kilo: Engineers need to be the CEOs of agents. Are they ready?Beyond MLOps: Building AI systems with MetaflowAgentic AI: From Risk Awareness to Practical Control | Noma SecurityShould agents be durable? | RenderGuardrails for the Future AI Safety and Responsible AI in PracticeAI Launchpad 2025: MooncakeThe Middle Ground: Balancing Batch and Real-Time Processing in a Data LakehouseAI Launchpad 2025: TopK
AI Council |

Optimizing Model Training End-to-End: A Tiny MoE Case Study Lambda

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER