
hashintel/hash🚀 The open-source, multi-tenant platform for self-building knowledge graphs and simulation
TypeScript · Rust★ 1,641↓ 244.7K/月2026年8月16日

Blaizzy/mlx-audioA text-to-speech (TTS), speech-to-text (STT) and speech-to-speech (STS) library built on Apple's MLX framework, providing efficient speech analysis on Apple Silicon.
Python★ 7,796↓ 582.2K/月2026年8月28日

Blaizzy/mlx-vlmMLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.
Python★ 5,431↓ 881.2K/月2026年8月28日
raullenchai/Rapid-MLXThe fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
Python★ 3,3922026年8月2日

defilantech/LLMKubeKubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 2072026年9月5日
drakulavich/kesha-voice-kitGive your tools a voice — speech to text and back, 25 languages, up to ~19× faster than Whisper. On your machine.
Rust · TypeScript★ 67↓ 1,029/月2026年7月28日

rajveer43/VeloxQuant-MLXFast KV-cache quantization for Apple Silicon (MLX) — 43 research-adapted compression methods with Metal kernels
Python★ 15↓ 7,549/月2026年9月5日
achiya-automation/safari-mcpNative Safari browser automation for AI agents. 80 tools via AppleScript — zero overhead, keeps logins, runs silently in background. Drop-in alternative to Chrome DevTools MCP with 40-60% less CPU/heat on Apple Silicon.
JavaScript★ 151↓ 6,053/月2026年7月17日
youssofal/MTPLX3x decode TPS increase On Qwen 3.6 27B @ temp 0.6 | Native MTP Speculative Decoding On Apple Silicon With No External Drafter.
Python · Swift★ 1,120↓ 2,694/月2026年8月2日
Python★ 83↓ 1,116/月2026年8月28日
Arthur-Ficial/apfelThe free AI already on your Mac. CLI tool, OpenAI-compatible server, and interactive chat — all on-device via Apple Intelligence. No API keys, no cloud, no downloads.
Swift · Python★ 6,2402026年8月4日

jundot/omlxLLM inference server with continuous batching & SSD caching for Apple Silicon — managed from the macOS menu bar
Python · Swift★ 18.5K2026年8月5日
Python★ 15↓ 2,711/月2026年7月18日
crates.io · npm · PyPI78良好健康指数

ohdearquant/latticeRun, quantize, and fine-tune LLMs on Apple Silicon. Pure Rust, no Python, no CUDA, no ONNX
Rust · Python★ 40↓ 19.1K/月2026年8月22日
michaelellis003/smcxState-space inference in JAX: Kalman and particle filters, tempered SMC, and SMC²
Python★ 32026年7月26日
tillahoffmann/jax-mpsA JAX backend for Apple Metal Performance Shaders (MPS), enabling GPU-accelerated JAX computations on Apple Silicon.
C++ · Python★ 192↓ 3,492/月2026年7月18日
arcships/light-ocrFast, offline OCR for Node.js & C++. PP-OCRv6 with Core ML / WebGPU hardware acceleration — recognize text in images with confidence scores & coordinates. npm: @arcships/light-ocr
C++ · JavaScript★ 4552026年7月31日
jagoff/memoPersistent semantic memory for AI agents — 100% local on Apple Silicon (MLX) or Linux/Ubuntu (CPU). Markdown source of truth, sqlite-vec + BM25 hybrid search, a codegraph-backed knowledge graph, MCP server + CLI. No cloud, no keys.
Python★ 7↓ 3,119/月2026年7月22日
binlecode/actopApple Silicon (M1–M4) power, GPU, ANE & memory-bandwidth monitor — sudoless TUI + Python API for profiling local LLM / MLX / CoreML inference
Python★ 0↓ 2,580/月2026年7月29日

asher/mlx-kquantNative K-quant support for MLX, with a quantization and fine-tuning toolchain for Apple Silicon
C++ · Python★ 6↓ 3,360/月2026年8月23日

ooples/AiDotNet.TensorsThe fastest .NET tensor library. Beats MathNet (6x), NumSharp (3200x), matches TorchSharp CPU - pure managed C# with hand-tuned AVX2/FMA SIMD kernels. Optional CUDA/OpenCL GPU acceleration.
C#★ 112026年9月6日
jjang-ai/vmlxvMLX - JANGTQ Uber Compressed MLX Models - L2 Disk Cache (survives restart) + L1 Paged (super fast ttft) + Hybrid SSM Scheduler + Cont Batching + etc!
Python · TypeScript★ 777↓ 4,850/月2026年7月24日
druide67/asiaiMulti-engine LLM benchmark & monitoring CLI for Apple Silicon
Python★ 11↓ 4,662/月2026年7月18日
PyPI · npm · RubyGems63中等健康指数
Ruby★ 1,360↓ 368/月2026年8月4日
DLTcollab/sse2neonA translator from Intel SSE intrinsics to Arm/Aarch64 NEON implementation
C++ · C · Python★ 1,5192026年7月20日
saiyam1814/kiacLocal Kubernetes on Apple's container framework - every node is its own lightweight VM. Metrics, storage, and LoadBalancer included.
Go★ 2612026年7月18日
lynicis/applecontainer-goA testcontainers-go-style Go library for spinning up Apple Container CLI Linux containers as test dependencies on macOS.
Go★ 32026年7月15日
jjang-ai/jangqJANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon
Python · Swift★ 2142026年7月22日

ARahim3/mlx-dsparkUp to 4× faster LLM decoding on Apple Silicon, lossless. Native MLX port of DeepSeek's DSpark & z-lab's DFlash speculative decoding — Gemma-4, Qwen3.8, Muse-Glimmer, Nemotron, Ornith-1.0, ternary Bonsai-27B.
Python · Swift★ 430↓ 5,564/月2026年8月19日
Python★ 0↓ 6,042/月2026年7月18日