On demand

Deploying LLMs on GKE with NVIDIA GPUs & Google Cloud

Hosted by Pulumi

Deploy a Mixtral 8x7B LLM on GKE with NVIDIA GPUs using Pulumi and Python. Learn to build scalable, GPU-enabled AI workloads on Google Cloud.

Watch the recording →
When
Watch any time
Format
On demand
Speakers
Engin Diri, Jason Smith
llmgpu infrastructurekubernetesgoogle cloudpython