Kubesploit: post #1865 — TG.ME

Forwarded fromKUKubeFM
Introducing Kube Signals: the new KubeFM show that turns keynote trends into direct conversations with the speakers shaping them.

For episode one, Brian Teller sits down with Saiyam Pathak from vCluster after his KubeCon India keynote on AI factories. They examine why the GPU beneath the model is becoming a platform-engineering problem.

They discuss:

- Why whole-GPU allocation wastes capacity
- How DRA, HAMI, MIG, and MPS enable sharing
- What Kubernetes must learn to support AI factories

Watch the full episode: https://ku.bz/4QZDqrnf-

This episode is sponsored by LearnKube. Download the free book, The Technical Guide to Kubernetes Rightsizing, to understand what Prometheus and Grafana cannot tell you about safely reducing requests and limits.
August 25, 2026 73 1