Operations | Monitoring | ITSM | DevOps | Cloud

17: There's No Life Without AI: Agents, MCP, and the Future of Automation With Viktor Farcic

On this episode of Kubex Talks, technology critic Viktor Farcic returns to talk with Andrew Hillier about the rapidly changing landscape in tech. Viktor has gone from AI skeptic to believer, claiming that there really is no life without AI anymore, from a professional standpoint.

Sharing GPUs without Flying Blind: Kubernetes Patterns for AI Inference

GPU sharing is quickly becoming a practical requirement for Kubernetes-based AI inference, as many modern workloads don’t need a full GPU to deliver value. But safely placing multiple containers on the same accelerator brings new challenges: scheduling, fairness, isolation, observability, and noisy-neighbor behavior. This 20 min session explore the GPU sharing landscape across Kubernetes: time-slicing, MPS, MIG, KAI Scheduler, and HAMi, and dives into the harder problem: operating shared GPUs in production, from tracking usage to enforcing fairness as demand shifts.

From Visibility to Real Savings: Turning FinOps Insights into Measurable Cost Reduction

FinOps programs are maturing, and most organizations have better visibility into cloud spend than ever before. Dashboards are full of data. And yet costs keep climbing. The problem isn’t the data. It’s the gap between knowing where the waste is and actually eliminating it. In this joint session, Tangoe and Kubex come together to bridge that gap. Tangoe brings deep expertise in spend management and FinOps discipline, while Kubex delivers infrastructure-level optimization across cloud, Kubernetes, and the AI and GPU workloads that are rapidly becoming the next frontier of cost pressure.

Autonomous K8s Optimization Involves Both Compute and Storage Resources - Are You Doing Both?

One of the most powerful capabilities in K8s is the ability to autoscale resources to meet demands, scaling resources up during peak periods to ensure performance, and down again during lower periods to save money. In this joint session, Lucidity and Kubex walk through what end-to-end K8s optimization looks like when you address both layers together. We cover: Expect real examples, not slides full of theory. You’ll leave with a clear picture of where waste is hiding in your environment and a prioritized approach to addressing it.

15: Optimizing AI Workloads: Balancing Cost, Performance, and Scalability with Bijit Ghosh

In this episode, Andrew Hillier and Bijit Ghosh discuss the evolving landscape of AI, discussing the growing prominence of inference over training, hybrid cloud strategies, balancing cost with performance, and the orchestration of complex hardware environments. The conversation also touches on emerging concepts like AI factories, the challenges of sovereign cloud, and how enterprises are navigating data gravity and regulatory constraints. It's a deep dive into optimizing AI infrastructure, managing costs, and the disruptive changes that are transforming both technology and business outcomes.