Uploaded August 2026 | Updated September 2026, 19 hours ago
Large Language Models require incredibly expensive hardware to run. As AI workloads become more complex—especially with long-context tasks like agentic coding that generate thousands of tokens per session—keeping costs down is a major challenge.
In this video short, discover how llm-d implements key performance optimizations to help you squeeze maximum token efficiency out of your GPUs.
Additional Resources:
What is llm-d and why do we need it? redhat.com/en/blog/what-llm-d-and-why-do-we-need-it
What is llm-d: redhat.com/en/topics/ai/what-is-llm-d
The future of AI should be open. The Red Hat Open Source and AI Program Office (OSAIPO) builds, champions, and sustains Red Hat's open source leadership and engagement in the AI era. We guide communities in the responsible integration of AI to accelerate innovation, increase collaboration, and shape open standards.
#opensource #artificialintelligence #machinelearning
Large Language Models require incredibly expensive hardware to run. As AI workloads become more complex—especially with long-context tasks like agentic coding that generate thousands of tokens per session—keeping costs down is a major challenge.
In this video short, discover how llm-d implements key performance optimizations to help you squeeze maximum token efficiency out of your GPUs.
Additional Resources:
What is llm-d and why do we need it? redhat.com/en/blog/what-llm-d-and-why-do-we-need-it
What is llm-d: redhat.com/en/topics/ai/what-is-llm-d
The future of AI should be open. The Red Hat Open Source and AI Program Office (OSAIPO) builds, champions, and sustains Red Hat's open source leadership and engagement in the AI era. We guide communities in the responsible integration of AI to accelerate innovation, increase collaboration, and shape open standards.
#opensource #artificialintelligence #machinelearning










