What Makes LLM Inference So Hard @WeightsBiases
What Makes LLM Inference So Hard  @WeightsBiases
Uploaded November 2025 | Updated September 2026, 2 weeks ago
Baseten CEO, Tuhin Srivastava explains why serious LLM workloads can’t rely on simple shared endpoints. Each model brings its own constraints, and the entire system has to adapt around them.

A real-time view into the complexity behind production-scale inference.

Full episode on the channel.
What Makes LLM Inference So HardA T cell foundation model for AI-powered target discovery and precision medicineIntroducing serverless reinforcement learning: Train reliable AI agents without worrying about GPUsBuild reliable AI agents using W&B TrainingThe FDA is slowing American companies?The Hidden Problem in Scaling LLM InferenceReal-time speech translation is changing everythingYou Can Log Videos to Your W&B RunsHow Runway Reached #1 in Video AIWhere AI Outworks HumansSynthetic data in medical device AI: Challenges and opportunitiesWhy Video Models Beat LLMs
Weights & Biases |

What Makes LLM Inference So Hard

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER