RamaLama: Making working with AI Models Boring by Cedric Clyburn @DevoxxForever
RamaLama: Making working with AI Models Boring by Cedric Clyburn  @DevoxxForever
Uploaded March 2026 | Updated September 2026, 2 weeks ago
#VoxxedDaysCERN26

Managing and deploying AI models can often require extensive system configuration and complex software dependencies. RamaLama, a new open-source tool, aims to make working with AI models straightforward by leveraging container technology, making the process "boring"—predictable, reliable, and easy to manage. RamaLama integrates with container engines like Podman and Docker to deploy AI models within containers, eliminating the need for manual configuration and ensuring optimal setup for both CPU and GPU systems.

This talk will introduce RamaLama’s key features, including support for multiple AI model registries (Ollama, Hugging Face, and OCI), simplified commands for running models as chatbots or REST API services, and compatibility with alternative AI runtimes like llama.cpp and vllm. We’ll explore RamaLama’s unique capabilities, such as generating Podman quadlet files for edge deployments and Kubernetes YAML for scalable deployment, demonstrating how it allows developers to transition from local experimentation to production seamlessly. Join us to learn how RamaLama enables frictionless, containerized AI model deployment for developers and system administrators alike.
RamaLama: Making working with AI Models Boring by Cedric ClyburnDatabase Connection Pool Sizing - Demystified! by Jasmin FluriDefense against the dark arts of AI: The playbook for a sovereign model-as-a-service platform[VDBUH2026] Magda Miu & Alin Miu - Engineering Leadership in the Age of AIThe Past, Present and Future of Programming Languages by Kevlin Henney - Voxxed Days CERN 26Running Your Coding Agent Locally: Lessons from a Real-World Experiment by S. Maestri and A. SoldanoWelcome to Voxxed Days Amsterdam 2026 - Opening (room 2)Pipeline Patterns and Antipatterns - Things your Pipeline Should (Not) Do by  Daniele Raniz RanelandHack Your Brain: Smarter Learning for Devs by Simone de Gijt[VDBUH2026] Abdel Sghiouar - Optimizing LLM Inference for the Rest of UsAccessibility powered by AI by Ramona DomenWho to blame? The AI, The Programmer, or The Prompt? by Makan Sepehrifar
Devoxx |

RamaLama: Making working with AI Models Boring by Cedric Clyburn

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER