GitHub Release Tracker
All JS React Ruby Go Postgres Frontend Node

llama

Past 7d, sorted by best first, all versions
20 results Markdown version
08/25 9
ollama/Ollama v0.33.0
Tool for running and managing large language models.
Go 179690☆ 1160d old #golang #llm #go #llama #llama2
08/28 7
goccy/go-llama v0.4.0
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
Go 152☆ 28d old #golang #ai #llm #go #llama
08/28 7
goccy/go-llama v0.3.0
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
Go 152☆ 28d old #golang #ai #llm #go #llama
08/28 7
ollama/Ollama v0.33.2
Tool for running and managing large language models.
Go 179690☆ 1160d old #golang #llm #go #llama #llama2
08/26 7
ollama/Ollama v0.33.1
Tool for running and managing large language models.
Go 179690☆ 1160d old #golang #llm #go #llama #llama2
08/25 7
lone-cloud/Gerbil v1.28.0
Desktop app for running Large Language Models locally with cross-platform support and integrated image generation.
TypeScript 472☆ 381d old #javascript #electron #typescript #llm #desktop
08/25 7
hybridgroup/yzma v1.25.0
Go-based library for hardware-accelerated local inference with llama.cpp integration.
Go 573☆ 342d old #golang #llm #go #llama #llamacpp
08/23 7
yoshoku/llama_cpp.rb v0.28.0
Ruby bindings for llama.cpp, enabling easy integration of the library in Ruby applications.
C 236☆ 1242d old #ruby #ai #llm #gem #c
08/27 6
helixml/HelixML 2.12.7
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
Go 803☆ 1047d old #golang #self-hosted #api #llm #openai
08/26 6
lone-cloud/Gerbil v1.28.1
Desktop app for running Large Language Models locally with cross-platform support and integrated image generation.
TypeScript 472☆ 381d old #javascript #electron #typescript #llm #desktop
08/24 6
helixml/HelixML 2.12.6
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
Go 803☆ 1047d old #golang #self-hosted #api #llm #openai
08/24 6
helixml/HelixML 2.12.5
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
Go 803☆ 1047d old #golang #self-hosted #api #llm #openai
08/27 5
goccy/go-llama v0.2.3
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
Go 152☆ 28d old #golang #ai #llm #go #llama
08/26 5
goccy/go-llama v0.2.2
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
Go 152☆ 28d old #golang #ai #llm #go #llama
08/26 5
tetherto/QVAC sdk-v0.18.2
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
TypeScript 556☆ 238d old #javascript #typescript #ai #llm #llama
08/25 5
goccy/go-llama v0.2.1
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
Go 152☆ 28d old #golang #ai #llm #go #llama
08/24 5
tetherto/QVAC sdk-v0.18.1
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
TypeScript 556☆ 238d old #javascript #typescript #ai #llm #llama
08/28 4
mybigday/llama.rn v0.13.0-rc.2
React Native binding for running LLaMA model inference with multimodal support including vision and audio.
C++ 1027☆ 1123d old #c++ #react-native #android #react #llm
08/23 4
mybigday/llama.rn v0.13.0-rc.1
React Native binding for running LLaMA model inference with multimodal support including vision and audio.
C++ 1027☆ 1123d old #c++ #react-native #android #react #llm
08/23 4
mostlygeek/llama-swap v251
Reliable on-demand model switching between local OpenAI-compatible inference servers without restarting applications.
Go 5504☆ 694d old #golang #go #llama #llamacpp #localllm