Усі теги
Тег каталогу

#gguf

Усі репозиторії публічного реєстру з цим тегом — із тем GitHub або ключових слів, які публікують їхні реєстри пакетів. Здоров'я вимірюється за тією ж версіонованою методологією, що й решта реєстру.

29 записів
З тегом «gguf»Упорядковано за індексом здоров'я
PyPI
99Винятковийіндекс здоров'я
ggml-org/llama.cpp
LLM inference in C/C++
C++ · C★ 122.7K4 серп. 2026 р.
MIT4 серп. 2026 р. · метрики 2.10.0
PyPI
97Винятковийіндекс здоров'я
intel/auto-round
A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support and full compatibility with vLLM, SGLang, and Transformers.
Python · C++★ 1 52016 лип. 2026 р.
Apache-2.016 лип. 2026 р. · метрики 2.10.0
npm
95Винятковийіндекс здоров'я
huggingface/huggingface.js
Use Hugging Face with JavaScript
TypeScript★ 2 503↓ 22.4M/міс27 серп. 2026 р.
MIT27 серп. 2026 р. · метрики 2.10.0
PyPI · npm
94Винятковийіндекс здоров'я
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K5 серп. 2026 р.
MIT5 серп. 2026 р. · метрики 2.10.0
Go
89Відміннийіндекс здоров'я
defilantech/LLMKube
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 2075 вер. 2026 р.
Apache-2.05 вер. 2026 р. · метрики 2.10.0
crates.io · PyPI · npm
87Відміннийіндекс здоров'я
AlexsJones/llmfit
Hundreds of models & providers. One command to find what runs on your hardware.
Rust★ 31.1K↓ 2 251/міс5 серп. 2026 р.
MIT5 серп. 2026 р. · метрики 2.10.0
npm
87Відміннийіндекс здоров'я
withcatai/node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 2 162↓ 3.6M/міс27 серп. 2026 р.
MIT27 серп. 2026 р. · метрики 2.10.0
PyPI
86Відміннийіндекс здоров'я
MakazhanAlpamys/Soup
Soup turns the pain of LLM fine-tuning into a simple workflow. One config, one command, done.
Python★ 7415 лип. 2026 р.
Apache-2.015 лип. 2026 р. · метрики 2.10.0
PyPI · npm
86Відміннийіндекс здоров'я
n24q02m/mcp-core
Shared foundation for building MCP servers -- Streamable HTTP transport, OAuth 2.1, browser-based credential setup, and a shared embedding daemon.
Python · TypeScript★ 1↓ 17.7K/міс22 серп. 2026 р.
Apache-2.022 серп. 2026 р. · метрики 2.10.0
crates.io · Maven
84Відміннийіндекс здоров'я
eugenehp/llama-cpp-rs
A wrapper around the llama-cpp library for rust, including new Sampler API from llama-cpp.
Rust★ 46↓ 7 448/міс22 серп. 2026 р.
Apache-2.022 серп. 2026 р. · метрики 2.10.0
Go
81Відміннийіндекс здоров'я
gpustack/gguf-parser-go
Review/Check GGUF files and estimate the memory usage and maximum tokens per second.
Go★ 2955 вер. 2026 р.
MIT5 вер. 2026 р. · метрики 2.10.0
Go · npm
80Відміннийіндекс здоров'я
kdeps/kdeps
Run AI workflows locally. Or deploy them anywhere. AI agent framework in YAML — workflow pipelines + autonomous agent loop. NVIDIA Inception member. Build, deploy, export as Docker/K8s/ISO.
Go★ 3524 лип. 2026 р.
Apache-2.024 лип. 2026 р. · метрики 2.10.0
PyPI
78Добрийіндекс здоров'я
Sahil170595/Chimeraforge
PyPI capacity-planning CLI for LLM deployment. pip install chimeraforge.
Python★ 2↓ 2 506/міс22 серп. 2026 р.
MIT22 серп. 2026 р. · метрики 2.10.0
crates.io
77Добрийіндекс здоров'я
ThreatFlux/gguf
A rust gguf library
Rust★ 6↓ 3 633/міс5 серп. 2026 р.
MIT5 серп. 2026 р. · метрики 2.10.0
Go
75Добрийіндекс здоров'я
hybridgroup/yzma
Go with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.
Go★ 52017 лип. 2026 р.
Власна ліцензія17 лип. 2026 р. · метрики 2.10.0
PyPI
71Добрийіндекс здоров'я
asher/mlx-kquant
Native K-quant support for MLX, with a quantization and fine-tuning toolchain for Apple Silicon
C++ · Python★ 6↓ 3 360/міс23 серп. 2026 р.
MIT23 серп. 2026 р. · метрики 2.10.0
PyPI · crates.io
67Добрийіндекс здоров'я
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7 559/міс17 лип. 2026 р.
MIT17 лип. 2026 р. · метрики 2.10.0
npm
67Добрийіндекс здоров'я
wundercorp/openmodel
Use any AI model, locally or in the cloud, while tracking on-device token usage and collecting token telemetry data.
TypeScript · JavaScript · CSS★ 18↓ 3 140/міс7 серп. 2026 р.
Apache-2.07 серп. 2026 р. · метрики 2.10.0
Go
65Добрийіндекс здоров'я
aimd54/palan
Pull, push, pack, and serve GGUF models as OCI ModelPack artifacts — daemonless, air-gap-first, one binary
Go★ 017 лип. 2026 р.
Apache-2.017 лип. 2026 р. · метрики 2.10.0
Go · PyPI
65Добрийіндекс здоров'я
anthony-chaudhary/fak
fak — the Fused Agent Kernel: one Go binary for AI agent loops. Wrap Claude Code/Codex/Cursor, keep long sessions cache-efficient, route work per call, run local GGUF models, and adjudicate tool calls.
Go · Python★ 1517 лип. 2026 р.
Apache-2.017 лип. 2026 р. · метрики 2.10.0
PyPI · crates.io
63Помірнийіндекс здоров'я
FedericoTs/quantprobe
Run a 110B on a 2016 PC with 16 GB RAM. Know your tok/s before you download. Placement beats budget: predicts speed + memory fit for any GGUF on your exact hardware, self-calibrates, emits the exact llama.cpp command — or 'quantprobe auto' does it all. Falsification-tested laws; misses published at full size. pip install quantprobe
Python★ 86↓ 6 494/міс20 серп. 2026 р.
MIT20 серп. 2026 р. · метрики 2.10.0
npm
62Помірнийіндекс здоров'я
therealtimex/node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 1↓ 41.2K/міс17 лип. 2026 р.
MIT17 лип. 2026 р. · метрики 2.10.0
60Помірнийіндекс здоров'я
eastriverlee/LLM.swift
LLM.swift is a simple and readable library that allows you to interact with large language models locally with ease for macOS, iOS, watchOS, tvOS, and visionOS.
Swift★ 86628 лип. 2026 р.
MIT28 лип. 2026 р. · метрики 2.10.0
Maven · npm
60Помірнийіндекс здоров'я
integrallis/models
In-JVM small language model inference
Java★ 34 вер. 2026 р.
Apache-2.04 вер. 2026 р. · метрики 2.10.0
npm
56Помірнийіндекс здоров'я
mohitsoni48/TurboLLM
Run any local LLM engine, auto-tuned to your GPU — polished web UI + OpenAI/Anthropic-compatible API. Point Claude Code at your own machine in one command. No Electron, no Python, offline-first.
TypeScript★ 173↓ 6 179/міс16 лип. 2026 р.
Без ліцензії16 лип. 2026 р. · метрики 2.10.0
PyPI
54Помірнийіндекс здоров'я
jjang-ai/jangq
JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon
Python · Swift★ 21422 лип. 2026 р.
Без ліцензії22 лип. 2026 р. · метрики 2.10.0
PyPI · npm · crates.io
53Помірнийіндекс здоров'я
JunHwan-Kwon/deepbom
Local static analysis and evidence generation for deployed AI model artifacts
JavaScript · Rust★ 2↓ 3 094/міс4 вер. 2026 р.
Apache-2.04 вер. 2026 р. · метрики 2.10.0
npm · crates.io · Maven
51Помірнийіндекс здоров'я
arusatech/llama-cpp-pro
Llama cpp + CapacitorJS support
Makefile · C · C++★ 7↓ 495/міс31 лип. 2026 р.
MIT31 лип. 2026 р. · метрики 2.10.0
44Слабкийіндекс здоров'я
calcuis/gguf-connector
gguf (GPT-Generated Unified Format) connector
Python★ 6031 серп. 2026 р.
MIT31 серп. 2026 р. · метрики 2.10.0