| 08/11 | 7 |
API for deploying and managing groups of pods as a single unit with leader and worker roles for multi-host inference workloads.
|
| 08/10 | 6 |
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
|
| 08/04 | 6 |
Kubernetes operator for enterprise-grade management and serving of large language models with model lifecycle automation, runtime selection, and GPU scheduling.
|
| 08/26 | 5 |
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
|
| 08/24 | 5 |
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
|
| 08/14 | 5 |
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
|
| 08/10 | 5 |
High-performance LLM proxy and load balancer providing intelligent routing, automatic failover, and unified model discovery across local and remote inference backends.
|
| 08/17 | 3 |
Pure Rust LLM inference engine focused on fast, hardware-optimized local model execution across GPUs and CPUs.
|
| 08/17 | 3 |
Pure Rust LLM inference engine focused on fast, hardware-optimized local model execution across GPUs and CPUs.
|
| 08/16 | 3 |
Pure Rust LLM inference engine focused on fast, hardware-optimized local model execution across GPUs and CPUs.
|
| 08/16 | 3 |
Pure Rust LLM inference engine focused on fast, hardware-optimized local model execution across GPUs and CPUs.
|
| 08/16 | 3 |
Pure Rust LLM inference engine focused on fast, hardware-optimized local model execution across GPUs and CPUs.
|
| 08/16 | 3 |
Pure Rust LLM inference engine focused on fast, hardware-optimized local model execution across GPUs and CPUs.
|