TGViewer
Kubesploit Kubesploit @kubesploit · 2.13K subscribers
Post #1294 372

Forwarded from KubeFM

John McBride, VP of Infrastructure and AI Engineering at the Linux Foundation shares how using Kubernetes and open-source AI models saved them tens of thousands of dollars.

You will learn:

- How to deploy VLLM on Kubernetes to serve open-source LLMs like Mistral and Llama
- How running inference workloads on your own infrastructure with T4 GPUs can reduce costs from tens of thousands to just a couple thousand dollars monthly
- Practical approaches to monitoring GPU workloads in production, including handling unpredictable failures and VRAM consumption issues

Watch (or listen to) it here: https://ku.bz/wP6bTlrFs

🌟 This episode is brought to you by StackGen! Don't let infrastructure block your teams. StackGen deterministically generates secure cloud infrastructure from any input - existing cloud environments, IaC or application code https://ku.bz/t0gBX9qQz

With @Birthmarkb "SWAG expert" Farrell
More from @kubesploit
  1. Oct 2, 2026Falco Event Generator creates suspicious system and Kubernetes activity so teams can safel…
  2. Oct 2, 2026Supply chain security is easier to reason about when it has layers. Meg Sarros frames cont…
  3. Oct 1, 2026Wardline is a self-hosted control-plane proxy for AI agents that enforces identity, policy…
  4. Sep 30, 2026FQDN Network Policy turns hostnames into current IP addresses and creates standard Kuberne…
  5. Sep 30, 2026"The next 10 years are going to be boring — in the good sense." Mauro Morales sees Kuberne…
  6. Sep 30, 2026This week on Learn Kubernetes Weekly 203: 🔥 Building Modelplane on Crossplane 🚪 Kubernet…
Threads Profile ViewerView any public Threads profile without an account.Open ThreadLook →Writing with AI? Make it sound human.Metric37 rewrites AI drafts so they read naturally. Free AI detector, 1,500 words free.Try Metric37 →