GitHub Release Tracker
All JS React Ruby Go Postgres Frontend Node

llama

Past 14d, sorted by best first, all versions
28 results Markdown version
08/25 9
ollama/Ollama v0.33.0
Tool for running and managing large language models.
Go 179558☆ 1158d old #golang #llm #go #llama #llama2
08/20 9
mudler/LocalAI v4.9.0
Self-hosted alternative to popular AI APIs for local inferencing on consumer-grade hardware.
Go 48702☆ 1257d old #golang #api #ai #llm #go
08/26 7
ollama/Ollama v0.33.1
Tool for running and managing large language models.
Go 179558☆ 1158d old #golang #llm #go #llama #llama2
08/25 7
lone-cloud/Gerbil v1.28.0
Desktop app for running Large Language Models locally with cross-platform support and integrated image generation.
TypeScript 472☆ 381d old #javascript #electron #typescript #llm #desktop
08/25 7
hybridgroup/yzma v1.25.0
Go-based library for hardware-accelerated local inference with llama.cpp integration.
Go 568☆ 339d old #golang #llm #go #llama #llamacpp
08/23 7
yoshoku/llama_cpp.rb v0.28.0
Ruby bindings for llama.cpp, enabling easy integration of the library in Ruby applications.
C 236☆ 1242d old #ruby #ai #llm #gem #c
08/21 7
hybridgroup/yzma v1.24.0
Go-based library for hardware-accelerated local inference with llama.cpp integration.
Go 568☆ 339d old #golang #llm #go #llama #llamacpp
08/19 7
ollama/Ollama v0.32.15
Tool for running and managing large language models.
Go 179558☆ 1158d old #golang #llm #go #llama #llama2
08/15 7
ollama/Ollama v0.32.14
Tool for running and managing large language models.
Go 179558☆ 1158d old #golang #llm #go #llama #llama2
08/14 7
ollama/Ollama v0.32.13
Tool for running and managing large language models.
Go 179558☆ 1158d old #golang #llm #go #llama #llama2
08/14 7
ollama/Ollama v0.32.12
Tool for running and managing large language models.
Go 179558☆ 1158d old #golang #llm #go #llama #llama2
08/13 7
ollama/Ollama v0.32.11
Tool for running and managing large language models.
Go 179558☆ 1158d old #golang #llm #go #llama #llama2
08/13 7
hybridgroup/yzma v1.23.0
Go-based library for hardware-accelerated local inference with llama.cpp integration.
Go 568☆ 339d old #golang #llm #go #llama #llamacpp
08/27 6
helixml/HelixML 2.12.7
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
Go 802☆ 1045d old #golang #self-hosted #api #llm #openai
08/26 6
lone-cloud/Gerbil v1.28.1
Desktop app for running Large Language Models locally with cross-platform support and integrated image generation.
TypeScript 472☆ 381d old #javascript #electron #typescript #llm #desktop
08/24 6
helixml/HelixML 2.12.6
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
Go 802☆ 1045d old #golang #self-hosted #api #llm #openai
08/24 6
helixml/HelixML 2.12.5
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
Go 802☆ 1045d old #golang #self-hosted #api #llm #openai
08/19 6
helixml/HelixML 2.12.4
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
Go 802☆ 1045d old #golang #self-hosted #api #llm #openai
08/13 6
helixml/HelixML 2.12.3
Private GenAI stack for deploying AI agents with support for RAG, API calls, vision, and efficient GPU scheduling.
Go 802☆ 1045d old #golang #self-hosted #api #llm #openai
08/27 5
goccy/go-llama v0.2.3
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
Go 149☆ 26d old #golang #ai #llm #go #llama
08/26 5
goccy/go-llama v0.2.2
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
Go 149☆ 26d old #golang #ai #llm #go #llama
08/26 5
tetherto/QVAC sdk-v0.18.2
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
TypeScript 554☆ 237d old #javascript #typescript #ai #llm #llama
08/25 5
goccy/go-llama v0.2.1
Pure Go inference engine for GGUF models, built from llama.cpp compiled to WebAssembly and translated to Go without wasm runtime.
Go 149☆ 26d old #golang #ai #llm #go #llama
08/24 5
tetherto/QVAC sdk-v0.18.1
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
TypeScript 554☆ 237d old #javascript #typescript #ai #llm #llama
08/14 5
tetherto/QVAC sdk-v0.17.1
Local-first, cross-platform SDK for building peer-to-peer AI apps with local model inference, speech, translation, and RAG.
TypeScript 554☆ 237d old #javascript #typescript #ai #llm #llama
08/23 4
mybigday/llama.rn v0.13.0-rc.1
React Native binding for running LLaMA model inference with multimodal support including vision and audio.
C++ 1026☆ 1121d old #c++ #react-native #android #react #llm
08/23 4
mostlygeek/llama-swap v251
Reliable on-demand model switching between local OpenAI-compatible inference servers without restarting applications.
Go 5485☆ 692d old #golang #go #llama #llamacpp #localllm
08/14 4
mostlygeek/llama-swap v250
Reliable on-demand model switching between local OpenAI-compatible inference servers without restarting applications.
Go 5485☆ 692d old #golang #go #llama #llamacpp #localllm