All
JS
React
Ruby
Go
Postgres
Frontend
Node
More...
javascript (13111)
typescript (8808)
go (6653)
golang (5956)
react (2171)
ruby (2071)
reactjs (1440)
nodejs (1104)
node (959)
ai (951)
cli (890)
hacktoberfest (658)
python (629)
css (623)
postgresql (602)
kubernetes (564)
rust (556)
docker (544)
postgres (529)
claude-code (515)
react-native (508)
electron (501)
ai-agents (484)
llm (364)
frontend (353)
vue (346)
claude (334)
html (334)
database (329)
android (322)
nextjs (307)
developer-tools (288)
api (283)
java (273)
mcp (267)
svelte (261)
agent (252)
rails (250)
automation (248)
linux (244)
ai-agent (240)
ios (232)
markdown (231)
chrome-extension (225)
security (214)
codex (207)
open-source (205)
macos (190)
agents (188)
php (186)
self-hosted (185)
c (184)
desktop-app (184)
angular (176)
mysql (174)
anthropic (166)
devops (165)
sql (158)
aws (156)
npm (152)
tailwindcss (151)
vite (151)
monitoring (146)
dashboard (144)
c++ (142)
cross-platform (142)
graphql (142)
framework (141)
bun (136)
github (136)
terraform (136)
blockchain (135)
browser (134)
git (133)
testing (130)
sdk (128)
http (123)
json (120)
ui (120)
editor (118)
agentic-ai (117)
github-actions (115)
deno (114)
chatgpt (113)
authentication (112)
expo (111)
ai-tools (108)
proxy (107)
web (107)
terminal (103)
components (102)
design-system (101)
containers (99)
analytics (98)
library (98)
mcp-server (98)
shell (97)
vue3 (97)
kotlin (96)
windows (96)
inference
Past 14d, sorted by best first
08/17
7
llm-d Router provides load and prefix-cache aware inference routing with prioritization and flow control, supporting standalone and Gateway API deployment via an endpoint picker. ✂
08/27
6
On-device mobile SDKs for running LLMs, speech-to-text, and text-to-speech locally. ✂
C++
10280☆
403d old
#swift
#c++
#llm
#ios
#kotlin
08/24
6
Platform for deploying, managing, and running GPU-accelerated inference, streaming, and batch workloads across worker clusters. ✂
08/20
4
Kubernetes operator managing self-hosted LLM inference on NVIDIA GPUs and Apple Silicon, with autoscaling, model routing, and OpenAI-compatible API. ✂
08/27
3
InferCrane provides an evidence-gated release lifecycle for self-hosted models behind a stable OpenAI-compatible endpoint, covering deploy, observe, scale, optimize, and safe promotion. ✂