* [NVIDIA GPU Operator v26.7.0](https://github.com/NVIDIA/gpu-operator) – Automates installation and lifecycle management of GPU drivers, container runtimes, and monitoring on Kubernetes nodes. * [ggrun v3.2.8](https://github.com/raketenkater/ggrun) – Auto-tuned launcher that measures multi-GPU hardware for GGUF models, picks an optimal llama.cpp/ik\_llama.cpp backend, and serves an OpenAI-compatible API. * [Cumo v0.5.10](https://github.com/sonots/cumo) – CUDA-aware GPU-optimized numerical library compatible with Ruby Numo for enhanced performance.