Cost-effective GPU workload optimisation on AWS using Kubernates and Spot Instances. Run AI workloads efficiently without breaking the bank!
-
Updated
Apr 26, 2025 - Python
Cost-effective GPU workload optimisation on AWS using Kubernates and Spot Instances. Run AI workloads efficiently without breaking the bank!
Kubernetes x NVIDIA DRA Workshop
This project provides a fully-integrated local AI (LLM) deployment, ready to bring up with a simple `./llm.sh` command. It handles various modes of operation.
To associate your repository with the nvidia-mig topic, visit your repo's landing page and select "manage topics."