On demand
Deploying LLMs on GKE with NVIDIA GPUs & Google Cloud
Hosted by Pulumi
Deploy a Mixtral 8x7B LLM on GKE with NVIDIA GPUs using Pulumi and Python. Learn to build scalable, GPU-enabled AI workloads on Google Cloud.
Watch the recording →- When
- Watch any time
- Format
- On demand
- Speakers
- Engin Diri, Jason Smith
llmgpu infrastructurekubernetesgoogle cloudpython