NVIDIA/gpu-operator resource
NVIDIA GPU Operator creates, configures, and manages GPUs in Kubernetes observed · 2026-08-28
Health v2 · maintenance only
95/100
- Activity 99
- Release rhythm 86
- Longevity 100
How is this computed?
round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.
- gap_med: 41
- age_days: 2745
- days_rel: 12
- days_push: 7
- n_releases_24m: 16
Adoption not part of the score
2851 stars · 534 forks observed · 2026-08-28
What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded
NVIDIA GPU Operator is a Kubernetes operator that automates the provisioning and lifecycle management of all NVIDIA software components needed to run GPUs in a cluster, including drivers, the device plugin, NVIDIA Container Toolkit, node labeling, and DCGM monitoring. It is deployed via Helm and lets administrators treat GPU nodes like standard CPU nodes using a common OS image.
Use cases
- automatically install nvidia gpu drivers on kubernetes nodes
- provision gpu worker nodes in a kubernetes cluster
- manage cuda drivers and container runtime in kubernetes
- monitor gpu usage in kubernetes with dcgm
- scale gpu nodes on cloud or on-prem kubernetes
- run nvidia vgpu in kubernetes
- allocate gpus to pods with device plugin or dra
When to choose
- you run Kubernetes clusters with NVIDIA GPUs and want automated driver and component lifecycle management
- you need to elastically scale GPU nodes without custom OS images
- you want DCGM-based GPU monitoring and node labeling out of the box
- you use OpenShift or NVIDIA AI Enterprise and need certified GPU provisioning
When to avoid
- you run GPUs outside Kubernetes or on single machines
- you have a tightly controlled environment where driver containers are not permitted
- you need non-NVIDIA GPU vendors
- your cluster nodes run unsupported OS versions or outdated kernels you cannot patch
Facets
infra-config · maturity active
container-orchestration deployment monitoring gpu-computing configuration-management cloud-computing gpu-computing machine-learning go cloud self-hosted kubernetes-operator nvidia cuda helm-chart device-plugin dcgm gpu-driver openshift vgpu dra containers devops kubernetes docker linux
10 sources
- readme: https://github.com/NVIDIA/gpu-operator · fetched 2026-08-28 · 6fa520ed92ee
- homepage: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/index.html · fetched 2026-08-29 · da306d64a36e
- site_page: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/overview.html · fetched 2026-08-29 · da306d64a36e
- site_page: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/getting-started.html · fetched 2026-08-29 · 81e710bfcd70
- site_page: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/uninstall.html · fetched 2026-08-29 · b48e4387c57c
- site_page: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/install-gpu-operator-vgpu.html · fetched 2026-08-29 · fedb6805dc5f
- site_page: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/install-gpu-operator-nvaie.html · fetched 2026-08-29 · dd38312cdc9f
- site_page: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/install-gpu-operator-gov-ready.html · fetched 2026-08-29 · b90f2965911a
- site_page: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/dra-intro-install.html · fetched 2026-08-29 · 0db8be32aeb5
- site_page: https://docs.nvidia.com/datacenter/cloud-native/gpu-operator/latest/install-gpu-operator-outdated-kernels.html · fetched 2026-08-29 · 3bb23e3af259
Member repositories
| Repository | Role | Health v2 |
|---|---|---|
| NVIDIA/gpu-operator | main | 95 |
For agents
markdown · JSON · MCP: product_card(name="NVIDIA/gpu-operator")
Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem