All tags
Catalogue tag

#local-llm

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

24 records
Tagged “local-llm”Ranked by health index
Go
74Goodhealth index
defilantech/llmkube
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 167↓ 0/moJul 14, 2026
Apache-2.0Jul 14, 2026 · metrics 1.13.0
PyPI
72Goodhealth index
MakazhanAlpamys/Soup
Soup turns the pain of LLM fine-tuning into a simple workflow. One config, one command, done.
Python★ 74Jul 15, 2026
Apache-2.0Jul 15, 2026 · metrics 1.13.0
npm
72Goodhealth index
gfargo/coco
AI-powered Git Assistant for CLI
TypeScript★ 12↓ 4,030/moJul 15, 2026
MITJul 15, 2026 · metrics 1.13.0
npm
70Goodhealth index
manojmallick/sigmap
97% token reduction for AI coding sessions — zero deps, 33 languages, MCP server
JavaScript★ 598↓ 10.8K/moJul 18, 2026
MITJul 18, 2026 · metrics 1.13.0
npm · Maven
69Moderatehealth index
mybigday/llama.rn
React Native binding of llama.cpp
C++ · C★ 1,000↓ 57.2K/moJul 17, 2026
MITJul 17, 2026 · metrics 1.13.0
npm
68Moderatehealth index
dcostenco/prism-coder
Persistent memory + local AI for coding agents. 2B–27B open-weight LLM fleet, cross-session Mind Palace, cognitive routing, L3 grounding verifier, multi-agent Hivemind. Works with Claude Code, Cursor, VS Code. Offline-first, HIPAA-ready. Free tier included.
TypeScript★ 155↓ 5,205/moJul 18, 2026
Apache-2.0Jul 18, 2026 · metrics 1.13.0
Go
67Moderatehealth index
ionalpha/flynn
A secure, self-improving agent operating system in a single Go binary. Bring any model, manage local models, point it at a goal, and grant it real authority: every action is sandboxed, governed, and sealed into a verifiable, tamper-evident record an independent party can check. Runs interactive or 24/7, or embed it in your own system.
Go★ 2Jul 18, 2026
Apache-2.0Jul 18, 2026 · metrics 1.13.0
npm
65Moderatehealth index
yuhp/opencode-models-discovery
A universal OpenCode plugin for dynamic model discovery with flexible configuration for OpenAI-compatible providers.
TypeScript★ 75↓ 15.9K/moJul 15, 2026
MITJul 15, 2026 · metrics 1.13.0
npm
63Moderatehealth index
m62624/pi-code-planner
Structured planning, bounded memory, TDD, and Git guardrails for local coding models in Pi Code (I don't know TypeScript at all; this is mostly a local-model experiment, with occasional help from Claude Code)
TypeScript★ 4↓ 3,041/moJul 22, 2026
MITJul 22, 2026 · metrics 1.13.0
Go
61Moderatehealth index
famclaw/famclaw
Self-hosted family AI gateway with parental controls. Runs on Linux, macOS, and Android (Termux) — Raspberry Pi, mini PC, old laptop, homelab server, even a phone. Telegram, Discord, web. Privacy-first, works with any LLM (local or cloud), OPA content filtering, MCP skill scanning.
Go★ 2Jul 17, 2026
AGPL-3.0Jul 17, 2026 · metrics 1.13.0
Go
60Moderatehealth index
rtmx-ai/aegis-cli
Air-gap-native agentic coding for closed environments — a hardened OpenCode TUI driven by a local model, with rtmx as the intent layer. Zero egress by construction.
Go★ 4Jul 20, 2026
Apache-2.0Jul 20, 2026 · metrics 1.13.0
Go · PyPI
59Moderatehealth index
anthony-chaudhary/fak
fak — the Fused Agent Kernel: one Go binary for AI agent loops. Wrap Claude Code/Codex/Cursor, keep long sessions cache-efficient, route work per call, run local GGUF models, and adjudicate tool calls.
Go · Python★ 15Jul 17, 2026
Apache-2.0Jul 17, 2026 · metrics 1.13.0
PyPI
57Moderatehealth index
moortekweb-art/agentic-harness
Self-hosted completion gate for coding agents: independent verification before done, durable evidence, CLI and local GUI.
Python★ 0↓ 2,023/moJul 18, 2026
MITJul 18, 2026 · metrics 1.13.0
crates.io · PyPI
56Moderatehealth index
ohdearquant/lattice
Run, quantize, and fine-tune LLMs on Apple Silicon. Pure Rust, no Python, no CUDA, no ONNX
Rust★ 31↓ 0/moJul 13, 2026
Apache-2.0Jul 13, 2026 · metrics 1.13.0
Go
55Moderatehealth index
mlhher/late-cli
Stop degrading your model's reasoning. A minimal, zero-config AI coding agent. Enforced ephemeral subagents keep context pure. From tiny local models up to Sol, Fable and Kimi K3.
Go★ 382Jul 20, 2026
Custom licenseJul 20, 2026 · metrics 1.13.0
npm
55Moderatehealth index
mohitsoni48/TurboLLM
Run any local LLM engine, auto-tuned to your GPU — polished web UI + OpenAI/Anthropic-compatible API. Point Claude Code at your own machine in one command. No Electron, no Python, offline-first.
TypeScript★ 173↓ 6,179/moJul 16, 2026
No licenseJul 16, 2026 · metrics 1.13.0
PyPI
52Moderatehealth index
JamesWHomer/jarv
Scriptable, multi-provider AI agent for the terminal — pipe into it, script it, run it on any model, cloud or local
Python★ 4↓ 8,378/moJul 14, 2026
MITJul 14, 2026 · metrics 1.13.0
PyPI
52Moderatehealth index
ToPo-ToPo-ToPo/local-llm-server
ローカルLLM(mlx / mlx-vlm / llama.cpp / router)を OpenAI 互換 API として起動・管理する軽量サーバー
Python★ 0↓ 4,302/moJul 19, 2026
Apache-2.0Jul 19, 2026 · metrics 1.13.0
Go · npm
51Moderatehealth index
liliang-cn/rago
AI Agent SDK designed for Go developers
Go★ 8Jul 18, 2026
MITJul 18, 2026 · metrics 1.13.0
npm
51Moderatehealth index
lloyal-ai/reasoning-run
A private reasoner for your terminal. Direct conversation or grounded multi-agent research, GPU-native and fully local. No API keys, no inference servers.
TypeScript · JavaScript★ 2↓ 2,264/moJul 19, 2026
MITJul 19, 2026 · metrics 1.13.0
PyPI
49At riskhealth index
Leon-Sander/Local-Multimodal-AI-Chat
Self-hostable multimodal chat with local LLMs (Ollama/OpenAI): PDF RAG, image chat, and Whisper voice, Streamlit + Docker.
Python★ 203Jul 18, 2026
GPL-3.0Jul 18, 2026 · metrics 1.13.0
npm
49At riskhealth index
eeshansrivastava89/offgrid-ai
Privacy-first CLI for running local LLMs — discover, configure, run, benchmark
JavaScript★ 1↓ 20.6K/moJul 17, 2026
No licenseJul 17, 2026 · metrics 1.13.0
PyPI · npm
38At riskhealth index
LearningCircuit/local-deep-research
~95% on SimpleQA (e.g. Qwen3.6-27B on a 3090). Supports all local and cloud LLMs (llama.cpp, Ollama, Google, ...). 10+ search engines - arXiv, PubMed, your private documents. Everything Local & Encrypted.
Python · JavaScript★ 8,752↓ 10.3K/moJul 20, 2026
MITJul 20, 2026 · metrics 1.13.0
npm
29Criticalhealth index
jamiejamesdev/pi-en2th
No repository description published.
TypeScript★ 0↓ 2,015/moJul 16, 2026
MITJul 16, 2026 · metrics 1.13.0