Livestream Mon 19 Jan
Infrastructure for AI: Managing GPU Workloads with Kubernetes
Hosted by O'Reilly
Saiyam Pathak speaks at O'Reilly's Infrastructure & Ops Superstream on managing GPU workloads at scale with Kubernetes. Covering GPU sharing, Kai Scheduler, multi-tenancy with vCluster, and scalable inference with vLLM.
View details →- When
- Mon 19 Jan
- Format
- Livestream
- Speaker
- Saiyam Pathak
- For
- infrastructure and ops practitioners
gpukubernetesinferencemulti-tenancy