Production-ready GPU-enabled Kubernetes AI infrastructure platform with NVIDIA GPU Operator, Ollama, CUDA workloads, autoscaling, monitoring, and distributed AI compute support.
- NVIDIA GPU Operator
- GPU-enabled Kubernetes workloads
- CUDA runtime support
- Ollama GPU inference
- GPU autoscaling
- Prometheus monitoring
- Grafana dashboards
- Node Feature Discovery
- NVIDIA Device Plugin
- Persistent storage
- Ingress support
- Production-ready manifests
- Kubernetes
- NVIDIA GPU Operator
- CUDA
- Ollama
- Docker
- Prometheus
- Grafana
- NGINX Ingress
- NVIDIA Container Toolkit
kubectl apply -f kubernetes/nexoryx-gpu- GPU Operator
- GPU Device Plugin
- Ollama GPU Pods
- Prometheus
- Grafana
- Ingress
- Autoscaling
- NVIDIA GPU Nodes
- NVIDIA Drivers
- Kubernetes Cluster
- Helm installed
- NVIDIA Container Toolkit
Update domains, storage classes, and GPU resource limits before production deployment.
- Kubernetes Helm charts
- GitOps support
- CI/CD improvements
- Monitoring dashboards
- Multi-cloud support
- Security hardening
This repository includes:
- Shell validation
- Markdown linting
- Terraform validation (where applicable)
See:
- examples/
- docs/
This repository is part of the Nexoryx infrastructure ecosystem.