Todas las etiquetas
Etiqueta del catálogo

#llama-cpp

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

27 registros
Con la etiqueta «llama-cpp»Ordenado por índice de salud
PyPI · npm
94Excepcionalíndice de salud
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K5 ago 2026
MIT5 ago 2026 · métricas 2.10.0
npm
93Excepcionalíndice de salud
Nano-Collective/nanocoder
An open coding agent for your terminal, built by a community collective rather than a company. Bring your own model, keep your code on your machine, and owe nothing to anyone.
TypeScript★ 2376↓ 9541/mes25 ago 2026
Licencia propia25 ago 2026 · métricas 2.10.0
npm
93Excepcionalíndice de salud
juspay/neurolink
One TypeScript interface for 24+ LLM providers — swap providers without rewriting. MCP-native (58+ servers), voice (TTS/STT/realtime), RAG, memory, file processors. Production-origin: powers Tara, Yama, and Clairvoyance at Juspay.
TypeScript★ 110↓ 24.2K/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
NuGet
89Excelenteíndice de salud
SciSharp/LLamaSharp
A C#/.NET library to run LLM (🦙LLaMA/LLaVA) on your local device efficiently.
C# · JavaScript★ 374218 jul 2026
MIT18 jul 2026 · métricas 2.10.0
Go
89Excelenteíndice de salud
defilantech/LLMKube
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 2075 sept 2026
Apache-2.05 sept 2026 · métricas 2.10.0
npm
87Excelenteíndice de salud
withcatai/node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 2162↓ 3.6M/mes27 ago 2026
MIT27 ago 2026 · métricas 2.10.0
Go
81Excelenteíndice de salud
gpustack/gguf-parser-go
Review/Check GGUF files and estimate the memory usage and maximum tokens per second.
Go★ 2955 sept 2026
MIT5 sept 2026 · métricas 2.10.0
npm · Maven
81Excelenteíndice de salud
mybigday/llama.rn
React Native binding of llama.cpp
C++ · C★ 1000↓ 57.2K/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
PyPI
78Buenoíndice de salud
Sahil170595/Chimeraforge
PyPI capacity-planning CLI for LLM deployment. pip install chimeraforge.
Python★ 2↓ 2506/mes22 ago 2026
MIT22 ago 2026 · métricas 2.10.0
npm
77Buenoíndice de salud
jcode-works/jcode-ragmir
Confidential local RAG for your coding agents.
TypeScript · JavaScript★ 6↓ 15.1K/mes26 jul 2026
AGPL-3.026 jul 2026 · métricas 2.10.0
npm
73Buenoíndice de salud
mybigday/llama.node
Node.js binding of llama.cpp
C++ · JavaScript · TypeScript★ 20↓ 1947/mes25 jul 2026
Sin licencia25 jul 2026 · métricas 2.10.0
Go
65Buenoíndice de salud
aimd54/palan
Pull, push, pack, and serve GGUF models as OCI ModelPack artifacts — daemonless, air-gap-first, one binary
Go★ 017 jul 2026
Apache-2.017 jul 2026 · métricas 2.10.0
PyPI · crates.io
63Moderadoíndice de salud
FedericoTs/quantprobe
Run a 110B on a 2016 PC with 16 GB RAM. Know your tok/s before you download. Placement beats budget: predicts speed + memory fit for any GGUF on your exact hardware, self-calibrates, emits the exact llama.cpp command — or 'quantprobe auto' does it all. Falsification-tested laws; misses published at full size. pip install quantprobe
Python★ 86↓ 6494/mes20 ago 2026
MIT20 ago 2026 · métricas 2.10.0
npm
63Moderadoíndice de salud
yoloshii/ClawMem
On-device memory layer for AI agents. Claude Code, Hermes and OpenClaw. Hooks + MCP server + hybrid RAG search.
TypeScript★ 194↓ 3556/mes25 jul 2026
MIT25 jul 2026 · métricas 2.10.0
npm
62Moderadoíndice de salud
IBazylchuk/paparats-mcp
Local‑first MCP server for multi‑repository semantic code search with Qdrant and llama. Turns your entire workspace into private context for AI coding assistants like Claude Code, Codex, Cursor, Copilot, Antigravity and Windsurf.
TypeScript★ 1018 jul 2026
MIT18 jul 2026 · métricas 2.10.0
Go
62Moderadoíndice de salud
tearingItUp786/chatgpt-tui
A portable terminal AI interface
Go★ 19717 jul 2026
MIT17 jul 2026 · métricas 2.10.0
Go
62Moderadoíndice de salud
tearingItUp786/nekot
A portable terminal AI interface
Go★ 19717 jul 2026
MIT17 jul 2026 · métricas 2.10.0
npm
62Moderadoíndice de salud
therealtimex/node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 1↓ 41.2K/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
npm
59Moderadoíndice de salud
egonSchiele/smoltalk
Wrapper for LLM APIs
TypeScript★ 0↓ 3307/mes4 ago 2026
Sin licencia4 ago 2026 · métricas 2.10.0
Go
57Moderadoíndice de salud
airiclenz/apogee
Terminal coding agent for local LLMs (llama.cpp, Ollama, vLLM) and any OpenAI-compatible API. OS-sandboxed autonomy, MCP, sessions. Go.
Go★ 421 ago 2026
MIT21 ago 2026 · métricas 2.10.0
npm
56Moderadoíndice de salud
mohitsoni48/TurboLLM
Run any local LLM engine, auto-tuned to your GPU — polished web UI + OpenAI/Anthropic-compatible API. Point Claude Code at your own machine in one command. No Electron, no Python, offline-first.
TypeScript★ 173↓ 6179/mes16 jul 2026
Sin licencia16 jul 2026 · métricas 2.10.0
PyPI
54Moderadoíndice de salud
danaug23/harness-arena
Your model, many harnesses, many benchmarks.
Python · HTML★ 2↓ 2988/mes23 ago 2026
Apache-2.023 ago 2026 · métricas 2.10.0
Go
54Moderadoíndice de salud
spencerbull/Yokai
Terminal-first GPU fleet manager for deploying, monitoring, and managing vLLM, llama.cpp, and ComfyUI workloads.
Go · TypeScript★ 116 ago 2026
MIT16 ago 2026 · métricas 2.10.0
npm · crates.io · Maven
51Moderadoíndice de salud
arusatech/llama-cpp-pro
Llama cpp + CapacitorJS support
Makefile · C · C++★ 7↓ 495/mes31 jul 2026
MIT31 jul 2026 · métricas 2.10.0
npm
51Moderadoíndice de salud
gsanhueza/pi-llama-cpp
Pi extension for llama.cpp integration
TypeScript★ 33↓ 9895/mes27 jul 2026
MIT27 jul 2026 · métricas 2.10.0
npm
50Moderadoíndice de salud
lloyal-ai/reasoning-run
A private reasoner for your terminal. Direct conversation or grounded multi-agent research, GPU-native and fully local. No API keys, no inference servers.
TypeScript · JavaScript★ 2↓ 2264/mes19 jul 2026
MIT19 jul 2026 · métricas 2.10.0
npm
48Débilíndice de salud
eeshansrivastava89/offgrid-ai
Privacy-first CLI for running local LLMs — discover, configure, run, benchmark
JavaScript★ 1↓ 20.6K/mes17 jul 2026
Sin licencia17 jul 2026 · métricas 2.10.0