Alle Tags
Katalog-Tag

#gguf

Alle Repositories im öffentlichen Register, die dieses Tag tragen — aus ihren GitHub-Topics oder den von ihren Paket-Registries veröffentlichten Schlagwörtern. Die Gesundheit wird nach derselben versionierten Methodik gemessen wie im übrigen Register.

29 Einträge
Getaggt als „gguf“Geordnet nach Gesundheitsindex
PyPI
99AußergewöhnlichGesundheitsindex
ggml-org/llama.cpp
LLM inference in C/C++
C++ · C★ 122.7K4. Aug. 2026
MIT4. Aug. 2026 · Metriken 2.10.0
PyPI
97AußergewöhnlichGesundheitsindex
intel/auto-round
A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support and full compatibility with vLLM, SGLang, and Transformers.
Python · C++★ 1.52016. Juli 2026
Apache-2.016. Juli 2026 · Metriken 2.10.0
npm
95AußergewöhnlichGesundheitsindex
huggingface/huggingface.js
Use Hugging Face with JavaScript
TypeScript★ 2.503↓ 22.4M/Monat27. Aug. 2026
MIT27. Aug. 2026 · Metriken 2.10.0
PyPI · npm
94AußergewöhnlichGesundheitsindex
modelscope/FunASR
Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
Python · C · HTML★ 19.6K5. Aug. 2026
MIT5. Aug. 2026 · Metriken 2.10.0
Go
89ExzellentGesundheitsindex
defilantech/LLMKube
Kubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 2075. Sept. 2026
Apache-2.05. Sept. 2026 · Metriken 2.10.0
crates.io · PyPI · npm
87ExzellentGesundheitsindex
AlexsJones/llmfit
Hundreds of models & providers. One command to find what runs on your hardware.
Rust★ 31.1K↓ 2.251/Monat5. Aug. 2026
MIT5. Aug. 2026 · Metriken 2.10.0
npm
87ExzellentGesundheitsindex
withcatai/node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 2.162↓ 3.6M/Monat27. Aug. 2026
MIT27. Aug. 2026 · Metriken 2.10.0
PyPI
86ExzellentGesundheitsindex
MakazhanAlpamys/Soup
Soup turns the pain of LLM fine-tuning into a simple workflow. One config, one command, done.
Python★ 7415. Juli 2026
Apache-2.015. Juli 2026 · Metriken 2.10.0
PyPI · npm
86ExzellentGesundheitsindex
n24q02m/mcp-core
Shared foundation for building MCP servers -- Streamable HTTP transport, OAuth 2.1, browser-based credential setup, and a shared embedding daemon.
Python · TypeScript★ 1↓ 17.7K/Monat22. Aug. 2026
Apache-2.022. Aug. 2026 · Metriken 2.10.0
crates.io · Maven
84ExzellentGesundheitsindex
eugenehp/llama-cpp-rs
A wrapper around the llama-cpp library for rust, including new Sampler API from llama-cpp.
Rust★ 46↓ 7.448/Monat22. Aug. 2026
Apache-2.022. Aug. 2026 · Metriken 2.10.0
Go
81ExzellentGesundheitsindex
gpustack/gguf-parser-go
Review/Check GGUF files and estimate the memory usage and maximum tokens per second.
Go★ 2955. Sept. 2026
MIT5. Sept. 2026 · Metriken 2.10.0
Go · npm
80ExzellentGesundheitsindex
kdeps/kdeps
Run AI workflows locally. Or deploy them anywhere. AI agent framework in YAML — workflow pipelines + autonomous agent loop. NVIDIA Inception member. Build, deploy, export as Docker/K8s/ISO.
Go★ 3524. Juli 2026
Apache-2.024. Juli 2026 · Metriken 2.10.0
PyPI
78GutGesundheitsindex
Sahil170595/Chimeraforge
PyPI capacity-planning CLI for LLM deployment. pip install chimeraforge.
Python★ 2↓ 2.506/Monat22. Aug. 2026
MIT22. Aug. 2026 · Metriken 2.10.0
crates.io
77GutGesundheitsindex
ThreatFlux/gguf
A rust gguf library
Rust★ 6↓ 3.633/Monat5. Aug. 2026
MIT5. Aug. 2026 · Metriken 2.10.0
Go
75GutGesundheitsindex
hybridgroup/yzma
Go with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.
Go★ 52017. Juli 2026
Eigene Lizenz17. Juli 2026 · Metriken 2.10.0
PyPI
71GutGesundheitsindex
asher/mlx-kquant
Native K-quant support for MLX, with a quantization and fine-tuning toolchain for Apple Silicon
C++ · Python★ 6↓ 3.360/Monat23. Aug. 2026
MIT23. Aug. 2026 · Metriken 2.10.0
PyPI · crates.io
67GutGesundheitsindex
handy-computer/transcribe.cpp
ggml speech-to-text inference for 16+ model families
C++ · C★ 197↓ 7.559/Monat17. Juli 2026
MIT17. Juli 2026 · Metriken 2.10.0
npm
67GutGesundheitsindex
wundercorp/openmodel
Use any AI model, locally or in the cloud, while tracking on-device token usage and collecting token telemetry data.
TypeScript · JavaScript · CSS★ 18↓ 3.140/Monat7. Aug. 2026
Apache-2.07. Aug. 2026 · Metriken 2.10.0
Go
65GutGesundheitsindex
aimd54/palan
Pull, push, pack, and serve GGUF models as OCI ModelPack artifacts — daemonless, air-gap-first, one binary
Go★ 017. Juli 2026
Apache-2.017. Juli 2026 · Metriken 2.10.0
Go · PyPI
65GutGesundheitsindex
anthony-chaudhary/fak
fak — the Fused Agent Kernel: one Go binary for AI agent loops. Wrap Claude Code/Codex/Cursor, keep long sessions cache-efficient, route work per call, run local GGUF models, and adjudicate tool calls.
Go · Python★ 1517. Juli 2026
Apache-2.017. Juli 2026 · Metriken 2.10.0
PyPI · crates.io
63MittelGesundheitsindex
FedericoTs/quantprobe
Run a 110B on a 2016 PC with 16 GB RAM. Know your tok/s before you download. Placement beats budget: predicts speed + memory fit for any GGUF on your exact hardware, self-calibrates, emits the exact llama.cpp command — or 'quantprobe auto' does it all. Falsification-tested laws; misses published at full size. pip install quantprobe
Python★ 86↓ 6.494/Monat20. Aug. 2026
MIT20. Aug. 2026 · Metriken 2.10.0
npm
62MittelGesundheitsindex
therealtimex/node-llama-cpp
Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 1↓ 41.2K/Monat17. Juli 2026
MIT17. Juli 2026 · Metriken 2.10.0
60MittelGesundheitsindex
eastriverlee/LLM.swift
LLM.swift is a simple and readable library that allows you to interact with large language models locally with ease for macOS, iOS, watchOS, tvOS, and visionOS.
Swift★ 86628. Juli 2026
MIT28. Juli 2026 · Metriken 2.10.0
Maven · npm
60MittelGesundheitsindex
integrallis/models
In-JVM small language model inference
Java★ 34. Sept. 2026
Apache-2.04. Sept. 2026 · Metriken 2.10.0
npm
56MittelGesundheitsindex
mohitsoni48/TurboLLM
Run any local LLM engine, auto-tuned to your GPU — polished web UI + OpenAI/Anthropic-compatible API. Point Claude Code at your own machine in one command. No Electron, no Python, offline-first.
TypeScript★ 173↓ 6.179/Monat16. Juli 2026
Keine Lizenz16. Juli 2026 · Metriken 2.10.0
PyPI
54MittelGesundheitsindex
jjang-ai/jangq
JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon
Python · Swift★ 21422. Juli 2026
Keine Lizenz22. Juli 2026 · Metriken 2.10.0
PyPI · npm · crates.io
53MittelGesundheitsindex
JunHwan-Kwon/deepbom
Local static analysis and evidence generation for deployed AI model artifacts
JavaScript · Rust★ 2↓ 3.094/Monat4. Sept. 2026
Apache-2.04. Sept. 2026 · Metriken 2.10.0
npm · crates.io · Maven
51MittelGesundheitsindex
arusatech/llama-cpp-pro
Llama cpp + CapacitorJS support
Makefile · C · C++★ 7↓ 495/Monat31. Juli 2026
MIT31. Juli 2026 · Metriken 2.10.0
44SchwachGesundheitsindex
calcuis/gguf-connector
gguf (GPT-Generated Unified Format) connector
Python★ 6031. Aug. 2026
MIT31. Aug. 2026 · Metriken 2.10.0