全部标签
目录标签

#local-llm

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

53 条记录
标签为“local-llm”按健康指数排序
PyPI · Maven
95卓越健康指数
maziyarpanahi/openmed
Local-first healthcare AI: clinical NER & HIPAA PII de-identification that runs 100% on-device. 2,200+ medical models, 21 languages, Apple MLX + Python, no cloud, no patient data leaving your network. Apache-2.0
Python★ 4,827↓ 4M/月2026年8月4日
Apache-2.02026年8月4日 · 指标 2.10.0
PyPI
93卓越健康指数
raullenchai/Rapid-MLX
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
Python★ 3,3922026年8月2日
Apache-2.02026年8月2日 · 指标 2.10.0
Go
89优秀健康指数
defilantech/LLMKube
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 2072026年9月5日
Apache-2.02026年9月5日 · 指标 2.10.0
PyPI
88优秀健康指数
Andyyyy64/whichllm
Find the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.
Python★ 6,4632026年8月24日
MIT2026年8月24日 · 指标 2.10.0
npm
87优秀健康指数
gfargo/coco
AI-powered Git Assistant for CLI
TypeScript★ 14↓ 3,502/月2026年9月6日
MIT2026年9月6日 · 指标 2.10.0
PyPI
86优秀健康指数
MakazhanAlpamys/Soup
Soup turns the pain of LLM fine-tuning into a simple workflow. One config, one command, done.
Python★ 742026年7月15日
Apache-2.02026年7月15日 · 指标 2.10.0
npm
83优秀健康指数
HybridAIOne/hybridclaw
Enterprise-ready self-hosted AI assistant runtime with sandboxed execution, secure credentials, approvals, and memory
TypeScript★ 126↓ 2,721/月2026年8月4日
MIT2026年8月4日 · 指标 2.10.0
PyPI
83优秀健康指数
skynetcmd/m3-memory
Local-first Memory Framework for AI Agents · 99.2% LongMemEval-S retrieval @ k=10 · Supports Claude · Antigravity · LangChain · Hermes · Gemini · OpenCode · OpenClaw · MCP-native and plugins · Hybrid search (FTS5 + vector + MMR) · GDPR · FIPS 140-3 ready · 100% local (fully offline) or cloud capable
Python★ 19↓ 8,178/月2026年8月9日
Apache-2.02026年8月9日 · 指标 2.10.0
PyPI
81优秀健康指数
jonigl/mcp-client-for-ollama
Harness the power of local LLMs with this TUI MCP Client for Ollama. Featuring all core MCP primitives (tools, prompts, resources), agent mode, multi-server, model switching, streaming responses, human-in-the-loop, thinking mode, model params config, system prompts, and saved preferences.
Python★ 782↓ 15.1K/月2026年7月24日
MIT2026年7月24日 · 指标 2.10.0
npm
81优秀健康指数
manojmallick/sigmap
97% token reduction for AI coding sessions — zero deps, 33 languages, MCP server
JavaScript★ 598↓ 10.8K/月2026年7月18日
MIT2026年7月18日 · 指标 2.10.0
npm · Maven
81优秀健康指数
mybigday/llama.rn
React Native binding of llama.cpp
C++ · C★ 1,000↓ 57.2K/月2026年7月17日
MIT2026年7月17日 · 指标 2.10.0
Go · npm · PyPI
81优秀健康指数
orneryd/NornicDB
Nornicdb is a distributed low-latency, Graph+Vector, Temporal MVCC with all sub-ms HNSW search, graph traversal, and writes. Using Neo4j Bolt/Cypher and qdrant's gRPC means you can switch with no changes while adding intelligent features like schemas, managed embeddings, reranking+llm, GPU accel, Auto-TLP, Policy-based Memory Decay, and MCP server.
Go★ 8332026年7月27日
MIT2026年7月27日 · 指标 2.10.0
npm
80优秀健康指数
dcostenco/prism-coder
Persistent memory + local AI for coding agents. 2B–27B open-weight LLM fleet, cross-session Mind Palace, cognitive routing, L3 grounding verifier, multi-agent Hivemind. Works with Claude Code, Cursor, VS Code. Offline-first, HIPAA-ready. Free tier included.
TypeScript★ 155↓ 5,205/月2026年7月18日
Apache-2.02026年7月18日 · 指标 2.10.0
Go
80优秀健康指数
ionalpha/flynn
A secure, self-improving agent operating system in a single Go binary. Bring any model, manage local models, point it at a goal, and grant it real authority: every action is sandboxed, governed, and sealed into a verifiable, tamper-evident record an independent party can check. Runs interactive or 24/7, or embed it in your own system.
Go★ 22026年7月18日
Apache-2.02026年7月18日 · 指标 2.10.0
crates.io · npm · PyPI
78良好健康指数
ohdearquant/lattice
Run, quantize, and fine-tune LLMs on Apple Silicon. Pure Rust, no Python, no CUDA, no ONNX
Rust · Python★ 40↓ 19.1K/月2026年8月22日
Apache-2.02026年8月22日 · 指标 2.10.0
PyPI
77良好健康指数
KevRojo/Dulus
Dulus Ai — Free Agentic AI, Making Gemini web cappable of running bash commands in your terminal! [Gui, Web, Cli, Telegram, 2,000 MCP, 100K Skills . LiteLLM (100+ providers), local models via Ollama, /lang in 34 languages, Mesa Redonda, I create the first utility coin that can be used 100% as AI quota or Fuel, is called $Dulus
Python★ 358↓ 7,957/月2026年9月5日
GPL-3.02026年9月5日 · 指标 2.10.0
npm
77良好健康指数
jcode-works/jcode-ragmir
Confidential local RAG for your coding agents.
TypeScript · JavaScript★ 6↓ 15.1K/月2026年7月26日
AGPL-3.02026年7月26日 · 指标 2.10.0
npm · crates.io
75良好健康指数
mlx-node/mlx-node
该仓库未发布描述。
Rust · TypeScript★ 152↓ 2,518/月2026年8月3日
MIT2026年8月3日 · 指标 2.10.0
PyPI
73良好健康指数
ASCIT31/Dark-Moon
Autonomous AI pentesting engine, continuous offensive security across web, cloud, identity, CI/CD, IaC, databases, Active Directory, Kubernetes and IoT firmware. Agentic reasoning plus real exploit execution deliver proof-based vulnerabilities. Privacy gateway: the LLM never sees your real IPs, hosts or creds, nothing leaves your perimeter.
Python · TypeScript · Shell★ 8022026年8月4日
GPL-3.02026年8月4日 · 指标 2.10.0
npm
73良好健康指数
mybigday/llama.node
Node.js binding of llama.cpp
C++ · JavaScript · TypeScript★ 20↓ 1,947/月2026年7月25日
无许可证2026年7月25日 · 指标 2.10.0
npm
73良好健康指数
yuhp/opencode-models-discovery
A universal OpenCode plugin for dynamic model discovery with flexible configuration for OpenAI-compatible providers.
TypeScript★ 75↓ 15.9K/月2026年7月15日
MIT2026年7月15日 · 指标 2.10.0
PyPI
71良好健康指数
asher/mlx-kquant
Native K-quant support for MLX, with a quantization and fine-tuning toolchain for Apple Silicon
C++ · Python★ 6↓ 3,360/月2026年8月23日
MIT2026年8月23日 · 指标 2.10.0
npm
71良好健康指数
itayinbarr/little-coder
A harness optimized to smaller LLMs
TypeScript · Python · JavaScript★ 1,830↓ 3,860/月2026年7月23日
Apache-2.02026年7月23日 · 指标 2.10.0
npm
71良好健康指数
m62624/pi-code-planner
Structured planning, bounded memory, TDD, and Git guardrails for local coding models in Pi Code (I don't know TypeScript at all; this is mostly a local-model experiment, with occasional help from Claude Code)
TypeScript★ 4↓ 3,041/月2026年7月22日
MIT2026年7月22日 · 指标 2.10.0
Go
69良好健康指数
famclaw/famclaw
Self-hosted family AI gateway with parental controls. Runs on Linux, macOS, and Android (Termux) — Raspberry Pi, mini PC, old laptop, homelab server, even a phone. Telegram, Discord, web. Privacy-first, works with any LLM (local or cloud), OPA content filtering, MCP skill scanning.
Go★ 22026年7月17日
AGPL-3.02026年7月17日 · 指标 2.10.0
Go
67良好健康指数
rtmx-ai/aegis-cli
Air-gap-native agentic coding for closed environments — a hardened OpenCode TUI driven by a local model, with rtmx as the intent layer. Zero egress by construction.
Go★ 42026年7月20日
Apache-2.02026年7月20日 · 指标 2.10.0
npm
67良好健康指数
wundercorp/openmodel
Use any AI model, locally or in the cloud, while tracking on-device token usage and collecting token telemetry data.
TypeScript · JavaScript · CSS★ 18↓ 3,140/月2026年8月7日
Apache-2.02026年8月7日 · 指标 2.10.0
Go · PyPI
65良好健康指数
anthony-chaudhary/fak
fak — the Fused Agent Kernel: one Go binary for AI agent loops. Wrap Claude Code/Codex/Cursor, keep long sessions cache-efficient, route work per call, run local GGUF models, and adjudicate tool calls.
Go · Python★ 152026年7月17日
Apache-2.02026年7月17日 · 指标 2.10.0
PyPI · crates.io
63中等健康指数
FedericoTs/quantprobe
Run a 110B on a 2016 PC with 16 GB RAM. Know your tok/s before you download. Placement beats budget: predicts speed + memory fit for any GGUF on your exact hardware, self-calibrates, emits the exact llama.cpp command — or 'quantprobe auto' does it all. Falsification-tested laws; misses published at full size. pip install quantprobe
Python★ 86↓ 6,494/月2026年8月20日
MIT2026年8月20日 · 指标 2.10.0
Go
62中等健康指数
shotah/ai-gantry
Personal AI agent you can actually own: one static Go binary, one persona, any OpenAI-compat LLM (Ollama, Gemini, Grok), MCP tools, chat via Telegram/Discord/Slack. Outbound-only — no dashboard, no config UI, no open ports, ever. Hardened so small local models actually finish tool calls.
Go★ 02026年8月11日
MIT2026年8月11日 · 指标 2.10.0