Todas las etiquetas
Etiqueta del catálogo

#local-llm

Todos los repositorios del registro público que llevan esta etiqueta, procedente de sus topics de GitHub o de las palabras clave que publican sus registros de paquetes. La salud se mide con la misma metodología versionada que el resto del registro.

53 registros
Con la etiqueta «local-llm»Ordenado por índice de salud
PyPI · Maven
95Excepcionalíndice de salud
maziyarpanahi/openmed
Local-first healthcare AI: clinical NER & HIPAA PII de-identification that runs 100% on-device. 2,200+ medical models, 21 languages, Apple MLX + Python, no cloud, no patient data leaving your network. Apache-2.0
Python★ 4827↓ 4M/mes4 ago 2026
Apache-2.04 ago 2026 · métricas 2.10.0
PyPI
93Excepcionalíndice de salud
raullenchai/Rapid-MLX
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
Python★ 33922 ago 2026
Apache-2.02 ago 2026 · métricas 2.10.0
Go
89Excelenteíndice de salud
defilantech/LLMKube
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 2075 sept 2026
Apache-2.05 sept 2026 · métricas 2.10.0
PyPI
88Excelenteíndice de salud
Andyyyy64/whichllm
Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.
Python★ 646324 ago 2026
MIT24 ago 2026 · métricas 2.10.0
npm
87Excelenteíndice de salud
gfargo/coco
AI-powered Git Assistant for CLI
TypeScript★ 14↓ 3502/mes6 sept 2026
MIT6 sept 2026 · métricas 2.10.0
PyPI
86Excelenteíndice de salud
MakazhanAlpamys/Soup
Soup turns the pain of LLM fine-tuning into a simple workflow. One config, one command, done.
Python★ 7415 jul 2026
Apache-2.015 jul 2026 · métricas 2.10.0
npm
83Excelenteíndice de salud
HybridAIOne/hybridclaw
Enterprise-ready self-hosted AI assistant runtime with sandboxed execution, secure credentials, approvals, and memory
TypeScript★ 126↓ 2721/mes4 ago 2026
MIT4 ago 2026 · métricas 2.10.0
PyPI
83Excelenteíndice de salud
skynetcmd/m3-memory
Local-first Memory Framework for AI Agents · 99.2% LongMemEval-S retrieval @ k=10 · Supports Claude · Antigravity · LangChain · Hermes · Gemini · OpenCode · OpenClaw · MCP-native and plugins · Hybrid search (FTS5 + vector + MMR) · GDPR · FIPS 140-3 ready · 100% local (fully offline) or cloud capable
Python★ 19↓ 8178/mes9 ago 2026
Apache-2.09 ago 2026 · métricas 2.10.0
PyPI
81Excelenteíndice de salud
jonigl/mcp-client-for-ollama
Harness the power of local LLMs with this TUI MCP Client for Ollama. Featuring all core MCP primitives (tools, prompts, resources), agent mode, multi-server, model switching, streaming responses, human-in-the-loop, thinking mode, model params config, system prompts, and saved preferences.
Python★ 782↓ 15.1K/mes24 jul 2026
MIT24 jul 2026 · métricas 2.10.0
npm
81Excelenteíndice de salud
manojmallick/sigmap
97% token reduction for AI coding sessions — zero deps, 33 languages, MCP server
JavaScript★ 598↓ 10.8K/mes18 jul 2026
MIT18 jul 2026 · métricas 2.10.0
npm · Maven
81Excelenteíndice de salud
mybigday/llama.rn
React Native binding of llama.cpp
C++ · C★ 1000↓ 57.2K/mes17 jul 2026
MIT17 jul 2026 · métricas 2.10.0
Go · npm · PyPI
81Excelenteíndice de salud
orneryd/NornicDB
Nornicdb is a distributed low-latency, Graph+Vector, Temporal MVCC with all sub-ms HNSW search, graph traversal, and writes. Using Neo4j Bolt/Cypher and qdrant's gRPC means you can switch with no changes while adding intelligent features like schemas, managed embeddings, reranking+llm, GPU accel, Auto-TLP, Policy-based Memory Decay, and MCP server.
Go★ 83327 jul 2026
MIT27 jul 2026 · métricas 2.10.0
npm
80Excelenteíndice de salud
dcostenco/prism-coder
Persistent memory + local AI for coding agents. 2B–27B open-weight LLM fleet, cross-session Mind Palace, cognitive routing, L3 grounding verifier, multi-agent Hivemind. Works with Claude Code, Cursor, VS Code. Offline-first, HIPAA-ready. Free tier included.
TypeScript★ 155↓ 5205/mes18 jul 2026
Apache-2.018 jul 2026 · métricas 2.10.0
Go
80Excelenteíndice de salud
ionalpha/flynn
A secure, self-improving agent operating system in a single Go binary. Bring any model, manage local models, point it at a goal, and grant it real authority: every action is sandboxed, governed, and sealed into a verifiable, tamper-evident record an independent party can check. Runs interactive or 24/7, or embed it in your own system.
Go★ 218 jul 2026
Apache-2.018 jul 2026 · métricas 2.10.0
crates.io · npm · PyPI
78Buenoíndice de salud
ohdearquant/lattice
Run, quantize, and fine-tune LLMs on Apple Silicon. Pure Rust, no Python, no CUDA, no ONNX
Rust · Python★ 40↓ 19.1K/mes22 ago 2026
Apache-2.022 ago 2026 · métricas 2.10.0
PyPI
77Buenoíndice de salud
KevRojo/Dulus
Dulus Ai — Free Agentic AI, Making Gemini web cappable of running bash commands in your terminal! [Gui, Web, Cli, Telegram, 2,000 MCP, 100K Skills . LiteLLM (100+ providers), local models via Ollama, /lang in 34 languages, Mesa Redonda, I create the first utility coin that can be used 100% as AI quota or Fuel, is called $Dulus
Python★ 358↓ 7957/mes5 sept 2026
GPL-3.05 sept 2026 · métricas 2.10.0
npm
77Buenoíndice de salud
jcode-works/jcode-ragmir
Confidential local RAG for your coding agents.
TypeScript · JavaScript★ 6↓ 15.1K/mes26 jul 2026
AGPL-3.026 jul 2026 · métricas 2.10.0
npm · crates.io
75Buenoíndice de salud
mlx-node/mlx-node
El repositorio no publica descripción.
Rust · TypeScript★ 152↓ 2518/mes3 ago 2026
MIT3 ago 2026 · métricas 2.10.0
PyPI
73Buenoíndice de salud
ASCIT31/Dark-Moon
Autonomous AI pentesting engine, continuous offensive security across web, cloud, identity, CI/CD, IaC, databases, Active Directory, Kubernetes and IoT firmware. Agentic reasoning plus real exploit execution deliver proof-based vulnerabilities. Privacy gateway: the LLM never sees your real IPs, hosts or creds, nothing leaves your perimeter.
Python · TypeScript · Shell★ 8024 ago 2026
GPL-3.04 ago 2026 · métricas 2.10.0
npm
73Buenoíndice de salud
mybigday/llama.node
Node.js binding of llama.cpp
C++ · JavaScript · TypeScript★ 20↓ 1947/mes25 jul 2026
Sin licencia25 jul 2026 · métricas 2.10.0
npm
73Buenoíndice de salud
yuhp/opencode-models-discovery
A universal OpenCode plugin for dynamic model discovery with flexible configuration for OpenAI-compatible providers.
TypeScript★ 75↓ 15.9K/mes15 jul 2026
MIT15 jul 2026 · métricas 2.10.0
PyPI
71Buenoíndice de salud
asher/mlx-kquant
Native K-quant support for MLX, with a quantization and fine-tuning toolchain for Apple Silicon
C++ · Python★ 6↓ 3360/mes23 ago 2026
MIT23 ago 2026 · métricas 2.10.0
npm
71Buenoíndice de salud
itayinbarr/little-coder
A harness optimized to smaller LLMs
TypeScript · Python · JavaScript★ 1830↓ 3860/mes23 jul 2026
Apache-2.023 jul 2026 · métricas 2.10.0
npm
71Buenoíndice de salud
m62624/pi-code-planner
Structured planning, bounded memory, TDD, and Git guardrails for local coding models in Pi Code (I don't know TypeScript at all; this is mostly a local-model experiment, with occasional help from Claude Code)
TypeScript★ 4↓ 3041/mes22 jul 2026
MIT22 jul 2026 · métricas 2.10.0
Go
69Buenoíndice de salud
famclaw/famclaw
Self-hosted family AI gateway with parental controls. Runs on Linux, macOS, and Android (Termux) — Raspberry Pi, mini PC, old laptop, homelab server, even a phone. Telegram, Discord, web. Privacy-first, works with any LLM (local or cloud), OPA content filtering, MCP skill scanning.
Go★ 217 jul 2026
AGPL-3.017 jul 2026 · métricas 2.10.0
Go
67Buenoíndice de salud
rtmx-ai/aegis-cli
Air-gap-native agentic coding for closed environments — a hardened OpenCode TUI driven by a local model, with rtmx as the intent layer. Zero egress by construction.
Go★ 420 jul 2026
Apache-2.020 jul 2026 · métricas 2.10.0
npm
67Buenoíndice de salud
wundercorp/openmodel
Use any AI model, locally or in the cloud, while tracking on-device token usage and collecting token telemetry data.
TypeScript · JavaScript · CSS★ 18↓ 3140/mes7 ago 2026
Apache-2.07 ago 2026 · métricas 2.10.0
Go · PyPI
65Buenoíndice de salud
anthony-chaudhary/fak
fak — the Fused Agent Kernel: one Go binary for AI agent loops. Wrap Claude Code/Codex/Cursor, keep long sessions cache-efficient, route work per call, run local GGUF models, and adjudicate tool calls.
Go · Python★ 1517 jul 2026
Apache-2.017 jul 2026 · métricas 2.10.0
PyPI · crates.io
63Moderadoíndice de salud
FedericoTs/quantprobe
Run a 110B on a 2016 PC with 16 GB RAM. Know your tok/s before you download. Placement beats budget: predicts speed + memory fit for any GGUF on your exact hardware, self-calibrates, emits the exact llama.cpp command — or 'quantprobe auto' does it all. Falsification-tested laws; misses published at full size. pip install quantprobe
Python★ 86↓ 6494/mes20 ago 2026
MIT20 ago 2026 · métricas 2.10.0
Go
62Moderadoíndice de salud
shotah/ai-gantry
Personal AI agent you can actually own: one static Go binary, one persona, any OpenAI-compat LLM (Ollama, Gemini, Grok), MCP tools, chat via Telegram/Discord/Slack. Outbound-only — no dashboard, no config UI, no open ports, ever. Hardened so small local models actually finish tool calls.
Go★ 011 ago 2026
MIT11 ago 2026 · métricas 2.10.0