Livestream Mon 19 Jan

Infrastructure for AI: Managing GPU Workloads with Kubernetes

Hosted by O'Reilly

Saiyam Pathak speaks at O'Reilly's Infrastructure & Ops Superstream on managing GPU workloads at scale with Kubernetes. Covering GPU sharing, Kai Scheduler, multi-tenancy with vCluster, and scalable inference with vLLM.

View details →
When
Mon 19 Jan
Format
Livestream
Speaker
Saiyam Pathak
For
infrastructure and ops practitioners
gpukubernetesinferencemulti-tenancy