OpsCanary
kubernetesai workloadsPractitioner

Transforming AI Workloads: My Journey from Attendee to Speaker at KubeCon India 2026

5 min read CNCF BlogSep 22, 2026Reviewed for accuracy
Share
PractitionerHands-on experience recommended

KubeCon + CloudNativeCon is more than just a conference; it’s a platform for sharing groundbreaking ideas and real-world solutions. My experience transitioning from an attendee to a speaker was fueled by a desire to tackle the complexities of AI workloads in Kubernetes. The challenge was clear: how do you efficiently manage and schedule GPU resources for AI tasks without losing control over your infrastructure?

In my talk, I detailed the process of transforming NVIDIA’s DGX Spark into a self-hosted AI cluster. We built a Kubernetes cluster on this powerful hardware, exposing GPUs to workloads effectively. The key to our success was employing Dynamic Resource Allocation (DRA) to schedule these resources properly. This approach not only maximized resource utilization but also ensured that we could serve models on infrastructure we fully owned and controlled, a critical aspect for organizations looking to maintain data sovereignty and operational efficiency.

As you consider implementing similar solutions, remember that the real-world application of these concepts can be complex. Understanding how to expose GPUs and manage workloads effectively is crucial. The insights shared at KubeCon are invaluable for anyone looking to push the boundaries of what Kubernetes can do in the realm of AI and machine learning.

Key takeaways

  • Leverage Dynamic Resource Allocation (DRA) to optimize GPU scheduling.
  • Build a Kubernetes cluster on NVIDIA's DGX Spark for robust AI workloads.
  • Control your infrastructure by serving models on self-hosted resources.

Why it matters

Efficiently managing AI workloads on Kubernetes can significantly enhance performance and resource utilization, leading to cost savings and improved operational control.

When NOT to use this

The official docs don't call out specific anti-patterns here. Use your judgment based on your scale and requirements.

Want the complete reference?

Read official docs

Test what you just learned

Quiz questions written from this article

Take the quiz →
Linux FoundationSponsor

Industry-standard certifications built by the people behind Linux and Kubernetes. Earn the CKA — the gold standard Kubernetes administrator cert. OpsCanary readers get 30% off year-round with code OPSCANARY3.

Get CKA certified →

Get the daily digest

One email. 5 articles. Every morning.

No spam. Unsubscribe anytime.