* [RunAnywhere v0.20.31](https://github.com/RunanywhereAI/runanywhere-sdks) – On-device mobile SDKs for running LLMs, speech-to-text, and text-to-speech locally. * [NVIDIA Cloud Functions (NVCF) deploy/helm/containe...](https://github.com/NVIDIA/nvcf) – Platform for deploying, managing, and running GPU-accelerated inference, streaming, and batch workloads across worker clusters. * [InferCrane v1.0.0-rc.1](https://github.com/infercrane/infercrane) – InferCrane provides an evidence-gated release lifecycle for self-hosted models behind a stable OpenAI-compatible endpoint, covering deploy, observe, scale, optimize, and safe promotion.