Alle Tags
Katalog-Tag

#inference

Alle Repositories im öffentlichen Register, die dieses Tag tragen — aus ihren GitHub-Topics oder den von ihren Paket-Registries veröffentlichten Schlagwörtern. Die Gesundheit wird nach derselben versionierten Methodik gemessen wie im übrigen Register.

106 Einträge
Getaggt als „inference“Geordnet nach Gesundheitsindex
PyPI · crates.io
99AußergewöhnlichGesundheitsindex
ai-dynamo/dynamo
A Datacenter Scale Distributed Inference Serving Framework
Rust · Python · Go★ 7.886↓ 59.4K/Monat28. Aug. 2026
Eigene Lizenz28. Aug. 2026 · Metriken 2.10.0
PyPI
99AußergewöhnlichGesundheitsindex
huggingface/transformers
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python★ 163.3K↓ 179M/Monat4. Aug. 2026
Apache-2.04. Aug. 2026 · Metriken 2.10.0
PyPI · crates.io
99AußergewöhnlichGesundheitsindex
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Python★ 88.2K4. Aug. 2026
Apache-2.04. Aug. 2026 · Metriken 2.10.0
PyPI · Go
98AußergewöhnlichGesundheitsindex
LMCache/LMCache
LMCache: Supercharge Your LLM with the Fastest KV Cache Layer
Python★ 11.2K↓ 98.8K/Monat15. Aug. 2026
Apache-2.015. Aug. 2026 · Metriken 2.10.0
PyPI · crates.io
98AußergewöhnlichGesundheitsindex
apache/tvm-ffi
Open ABI and FFI for Machine Learning Systems
C++ · Python · Rust★ 452↓ 7.9M/Monat27. Aug. 2026
Apache-2.027. Aug. 2026 · Metriken 2.10.0
PyPI
98AußergewöhnlichGesundheitsindex
kvcache-ai/Mooncake
Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
C++ · Python★ 6.25612. Aug. 2026
Apache-2.012. Aug. 2026 · Metriken 2.10.0
PyPI · crates.io
98AußergewöhnlichGesundheitsindex
sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
Python★ 31.3K↓ 144M/Monat4. Aug. 2026
Apache-2.04. Aug. 2026 · Metriken 2.10.0
PyPI
98AußergewöhnlichGesundheitsindex
triton-inference-server/server
The Triton Inference Server provides an optimized cloud and edge inferencing solution.
Python · Shell · C++★ 10.9K27. Aug. 2026
BSD-3-Clause27. Aug. 2026 · Metriken 2.10.0
PyPI · RubyGems
97AußergewöhnlichGesundheitsindex
deepspeedai/DeepSpeed
DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.
Python · C++★ 42.9K↓ 1.1M/Monat13. Aug. 2026
Apache-2.013. Aug. 2026 · Metriken 2.10.0
PyPI
97AußergewöhnlichGesundheitsindex
pytorch/ao
PyTorch native quantization and sparsity for training and inference
Python · C++★ 2.90921. Juli 2026
Eigene Lizenz21. Juli 2026 · Metriken 2.10.0
PyPI
96AußergewöhnlichGesundheitsindex
mozilla-ai/any-llm
Communicate with an LLM provider using a single interface
Python★ 2.133↓ 130.6K/Monat21. Juli 2026
Apache-2.021. Juli 2026 · Metriken 2.10.0
npm
96AußergewöhnlichGesundheitsindex
vercel/ai
The AI Toolkit for TypeScript. From the creators of Next.js, the AI SDK is a free open-source library for building AI-powered applications and agents
TypeScript · MDX★ 26K↓ 87M/Monat5. Aug. 2026
Eigene Lizenz5. Aug. 2026 · Metriken 2.10.0
PyPI · crates.io
95AußergewöhnlichGesundheitsindex
ai-dynamo/aiconfigurator
Offline optimization of your disaggregated Dynamo graph
Python · Rust★ 37426. Juli 2026
Apache-2.026. Juli 2026 · Metriken 2.10.0
crates.io
95AußergewöhnlichGesundheitsindex
ai-dynamo/modelexpress
Model Express is a Rust-based component meant to be placed next to existing model inference systems to speed up their startup times and improve overall performance.
Python · Rust★ 103↓ 246.6K/Monat1. Aug. 2026
Apache-2.01. Aug. 2026 · Metriken 2.10.0
PyPI
95AußergewöhnlichGesundheitsindex
gpustack/gpustack
A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.
Python★ 5.422↓ 2.132/Monat2. Aug. 2026
Apache-2.02. Aug. 2026 · Metriken 2.10.0
npm
95AußergewöhnlichGesundheitsindex
huggingface/huggingface.js
Use Hugging Face with JavaScript
TypeScript★ 2.503↓ 22.4M/Monat27. Aug. 2026
MIT27. Aug. 2026 · Metriken 2.10.0
PyPI
94AußergewöhnlichGesundheitsindex
huggingface/optimum
🚀 Accelerate inference and training of 🤗 Transformers, Diffusers, TIMM and Sentence Transformers with easy to use hardware optimization tools
Python★ 3.44821. Juli 2026
Apache-2.021. Juli 2026 · Metriken 2.10.0
crates.io · npm · PyPI
94AußergewöhnlichGesundheitsindex
pykeio/ort
Fast ML inference & training for ONNX models in Rust
Rust★ 2.477↓ 4.3M/Monat27. Aug. 2026
Apache-2.027. Aug. 2026 · Metriken 2.10.0
npm · RubyGems
94AußergewöhnlichGesundheitsindex
vega/vega
A visualization grammar.
JavaScript · TypeScript★ 12K↓ 28.4M/Monat27. Aug. 2026
BSD-3-Clause27. Aug. 2026 · Metriken 2.10.0
PyPI · npm · RubyGems
93AußergewöhnlichGesundheitsindex
mlc-ai/xgrammar
Fast, Flexible and Portable Structured Generation
C++ · Python★ 1.845↓ 8.2M/Monat27. Aug. 2026
Apache-2.027. Aug. 2026 · Metriken 2.10.0
PyPI
93AußergewöhnlichGesundheitsindex
pytorch/TensorRT
PyTorch/TorchScript/FX compiler for NVIDIA GPUs using TensorRT
Python · Jupyter Notebook · C++★ 2.98430. Juli 2026
BSD-3-Clause30. Juli 2026 · Metriken 2.10.0
PyPI
93AußergewöhnlichGesundheitsindex
raullenchai/Rapid-MLX
The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
Python★ 3.3922. Aug. 2026
Apache-2.02. Aug. 2026 · Metriken 2.10.0
Go
93AußergewöhnlichGesundheitsindex
run-ai/karta
Translation layer that maps any Kubernetes framework's Custom Resource Definitions (CRDs) into a standardized, generic structure.
Go★ 605. Aug. 2026
Apache-2.05. Aug. 2026 · Metriken 2.10.0
crates.io · PyPI
92ExzellentGesundheitsindex
lightseekorg/smg
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing, chat history, tokenization caching, Responses API, embeddings, WASM plugins, MCP, and multi-tenant auth.
Rust · Python★ 398↓ 39/Monat16. Juli 2026
Apache-2.016. Juli 2026 · Metriken 2.10.0
PyPI · npm
92ExzellentGesundheitsindex
roboflow/inference
Turn any computer or edge device into a command center for your computer vision projects.
Python★ 2.4385. Sept. 2026
Eigene Lizenz5. Sept. 2026 · Metriken 2.10.0
PyPI
91ExzellentGesundheitsindex
triton-inference-server/model_analyzer
Triton Model Analyzer is a CLI tool to help with better understanding of the compute and memory requirements of the Triton Inference Server models.
Python★ 523↓ 10.6K/Monat29. Juli 2026
Apache-2.029. Juli 2026 · Metriken 2.10.0
Go
91ExzellentGesundheitsindex
utkuozdemir/nvidia_gpu_exporter
Nvidia GPU exporter for prometheus using nvidia-smi binary
Go★ 1.51022. Juli 2026
MIT22. Juli 2026 · Metriken 2.10.0
90ExzellentGesundheitsindex
NVIDIA/nvcf
Platform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.
Go · Rust★ 18017. Juli 2026
Apache-2.017. Juli 2026 · Metriken 2.10.0
PyPI
90ExzellentGesundheitsindex
Tencent/ncnn
ncnn is a high-performance neural network inference framework optimized for the mobile platform
C++ · C★ 23.7K12. Aug. 2026
Eigene Lizenz12. Aug. 2026 · Metriken 2.10.0
npm
90ExzellentGesundheitsindex
colinhacks/zod
TypeScript-first schema validation with static type inference
TypeScript★ 43.4K↓ 992M/Monat4. Aug. 2026
MIT4. Aug. 2026 · Metriken 2.10.0