All tags
Catalogue tag

#sglang

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

12 records
Tagged “sglang”Ranked by health index
PyPI · crates.io
99Exceptionalhealth index
ai-dynamo/dynamo
A Datacenter Scale Distributed Inference Serving Framework
Rust · Python · Go★ 7,886↓ 59.4K/moAug 28, 2026
Custom licenseAug 28, 2026 · metrics 2.10.0
PyPI
98Exceptionalhealth index
kvcache-ai/Mooncake
Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
C++ · Python★ 6,256Aug 12, 2026
Apache-2.0Aug 12, 2026 · metrics 2.10.0
PyPI
97Exceptionalhealth index
intel/auto-round
A SOTA quantization algorithm for high-accuracy low-bit LLM inference, seamlessly optimized for CPU/XPU/CUDA, with multi-datatype support and full compatibility with vLLM, SGLang, and Transformers.
Python · C++★ 1,520Jul 16, 2026
Apache-2.0Jul 16, 2026 · metrics 2.10.0
PyPI
95Exceptionalhealth index
gpustack/gpustack
A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.
Python★ 5,422↓ 2,132/moAug 2, 2026
Apache-2.0Aug 2, 2026 · metrics 2.10.0
crates.io · PyPI
92Excellenthealth index
lightseekorg/smg
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
Rust · Python★ 398↓ 39/moJul 16, 2026
Apache-2.0Jul 16, 2026 · metrics 2.10.0
crates.io · PyPI
91Excellenthealth index
smg-project/smg
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
Rust · Python★ 432Aug 2, 2026
Apache-2.0Aug 2, 2026 · metrics 2.10.0
PyPI
89Excellenthealth index
ModelCloud/GPTQModel
LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.
Python · Cuda★ 1,207Jul 17, 2026
Custom licenseJul 17, 2026 · metrics 2.10.0
Go · npm
89Excellenthealth index
matrixhub-ai/matrixhub
An Open-source, self-hosted AI model hub with Hugging Face compatibility, accelerating vLLM/SGLang performance.
Go · TypeScript★ 256Jul 17, 2026
Apache-2.0Jul 17, 2026 · metrics 2.10.0
Go · npm
87Excellenthealth index
ome-projects/ome
Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton
Go★ 481Jul 21, 2026
Apache-2.0Jul 21, 2026 · metrics 2.10.0
npm
77Goodhealth index
0xsero/vllm-studio
Control panel for VLLM, Sglang, llama.cpp, exllamav3
TypeScript★ 1,368Jul 16, 2026
Apache-2.0Jul 16, 2026 · metrics 2.10.0
npm
71Goodhealth index
happyvertical/ocr
No repository description published.
TypeScript★ 0↓ 3,363/moJul 15, 2026
MITJul 15, 2026 · metrics 2.10.0
PyPI · crates.io
62Moderatehealth index
theoddden/Terradev
An imperative command-line-interface for AI workload orchestration
Python★ 25↓ 3,845/moAug 23, 2026
Apache-2.0Aug 23, 2026 · metrics 2.10.0