| 08/25 | 9 |
Tool for running and managing large language models.
|
| 08/20 | 9 |
Self-hosted alternative to popular AI APIs for local inferencing on consumer-grade hardware.
|
| 08/25 | 7 |
Desktop app for running Large Language Models locally with cross-platform support and integrated image generation.
|
| 08/25 | 7 |
Go-based library for hardware-accelerated local inference with llama.cpp integration.
|
| 08/23 | 7 |
Ruby bindings for llama.cpp, enabling easy integration of the library in Ruby applications.
|
| 08/24 | 6 |
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
|
| 08/27 | 5 |
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
|
| 08/26 | 5 |
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
|
| 08/23 | 4 |
React Native binding for running LLaMA model inference with multimodal support including vision and audio.
|
| 08/23 | 4 |
Reliable on-demand model switching between local OpenAI-compatible inference servers without restarting applications.
|