* [NVIDIA AI Cluster Runtime v1.0.0](https://github.com/NVIDIA/aicr) – Tooling for optimized, validated, reproducible GPU-accelerated Kubernetes clusters via version-locked recipes and deployment-ready bundles. * [NVSentinel v1.26.0](https://github.com/NVIDIA/NVSentinel) – Cross-platform remediation service detecting, classifying, and automatically resolving runtime GPU node faults in Kubernetes clusters. * [Bend v2.0.36](https://github.com/bendlang/bend) – Fast language focused on enforcing application laws through formal proofs and compiling to high-performance executables. * [TypeGPU v0.12.7](https://github.com/software-mansion/TypeGPU) – TypeScript library enhancing the WebGPU API for type-safe resource management. * [GPUd v0.13.4](https://github.com/leptonai/gpud) – GPU-focused monitoring and diagnostics tool that detects GPU and fabric errors and reports critical system metrics. * [NVIDIA Cloud Functions (NVCF) deploy/helm/nvca-ope...](https://github.com/NVIDIA/nvcf) – Platform for deploying, managing, and running GPU-accelerated inference, streaming, and batch workloads across worker clusters. * [I3K RAG Engine v0.1.48](https://github.com/I3K-IT/RAG-Enterprise) – Self-hosted retrieval-augmented generation engine that keeps all document processing and answering offline on a single binary, with cited sources. * [Grove v0.1.0-alpha.14](https://github.com/ai-dynamo/grove) – Kubernetes API providing a single declarative interface to orchestrate multi-node AI inference with topology-aware placement, hierarchical gang scheduling, and autoscaling.