All tags
Catalogue tag

#llama-cpp

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

14 records
Tagged “llama-cpp”Ranked by health index
npm
78Goodhealth index
juspay/neurolink
One TypeScript interface for 24+ LLM providers — swap providers without rewriting. MCP-native (58+ servers), voice (TTS/STT/realtime), RAG, memory, file processors. Production-origin: powers Tara, Yama, and Clairvoyance at Juspay.
TypeScript★ 110↓ 24.2K/moJul 17, 2026
MITJul 17, 2026 · metrics 1.13.0
NuGet
75Goodhealth index
SciSharp/LLamaSharp
A C#/.NET library to run LLM (🦙LLaMA/LLaVA) on your local device efficiently.
C# · JavaScript★ 3,742Jul 18, 2026
MITJul 18, 2026 · metrics 1.13.0
Go
74Goodhealth index
defilantech/llmkube
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 167↓ 0/moJul 14, 2026
Apache-2.0Jul 14, 2026 · metrics 1.13.0
npm · Maven
69Moderatehealth index
mybigday/llama.rn
React Native binding of llama.cpp
C++ · C★ 1,000↓ 57.2K/moJul 17, 2026
MITJul 17, 2026 · metrics 1.13.0
Go
62Moderatehealth index
gpustack/gguf-parser-go
Review/Check GGUF files and estimate the memory usage and maximum tokens per second.
Go★ 276↓ 0/moJul 14, 2026
MITJul 14, 2026 · metrics 1.13.0
Go
60Moderatehealth index
aimd54/palan
Pull, push, pack, and serve GGUF models as OCI ModelPack artifacts — daemonless, air-gap-first, one binary
Go★ 0Jul 17, 2026
Apache-2.0Jul 17, 2026 · metrics 1.13.0
Go
59Moderatehealth index
tearingItUp786/chatgpt-tui
A portable terminal AI interface
Go★ 197Jul 17, 2026
MITJul 17, 2026 · metrics 1.13.0
Go
59Moderatehealth index
tearingItUp786/nekot
A portable terminal AI interface
Go★ 197Jul 17, 2026
MITJul 17, 2026 · metrics 1.13.0
npm
59Moderatehealth index
therealtimex/node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 1↓ 41.2K/moJul 17, 2026
MITJul 17, 2026 · metrics 1.13.0
npm
57Moderatehealth index
IBazylchuk/paparats-mcp
Local‑first MCP server for multi‑repository semantic code search with Qdrant and llama. Turns your entire workspace into private context for AI coding assistants like Claude Code, Codex, Cursor, Copilot, Antigravity and Windsurf.
TypeScript★ 10Jul 18, 2026
MITJul 18, 2026 · metrics 1.13.0
Go
57Moderatehealth index
thxcode/gguf-parser-go
Review/Check GGUF files and estimate the memory usage and maximum tokens per second.
Go★ 276Jul 15, 2026
MITJul 15, 2026 · metrics 1.13.0
npm
55Moderatehealth index
mohitsoni48/TurboLLM
Run any local LLM engine, auto-tuned to your GPU — polished web UI + OpenAI/Anthropic-compatible API. Point Claude Code at your own machine in one command. No Electron, no Python, offline-first.
TypeScript★ 173↓ 6,179/moJul 16, 2026
No licenseJul 16, 2026 · metrics 1.13.0
npm
51Moderatehealth index
lloyal-ai/reasoning-run
A private reasoner for your terminal. Direct conversation or grounded multi-agent research, GPU-native and fully local. No API keys, no inference servers.
TypeScript · JavaScript★ 2↓ 2,264/moJul 19, 2026
MITJul 19, 2026 · metrics 1.13.0
npm
49At riskhealth index
eeshansrivastava89/offgrid-ai
Privacy-first CLI for running local LLMs — discover, configure, run, benchmark
JavaScript★ 1↓ 20.6K/moJul 17, 2026
No licenseJul 17, 2026 · metrics 1.13.0