Усі теги
Тег каталогу

#model-serving

Усі репозиторії публічного реєстру з цим тегом — із тем GitHub або ключових слів, які публікують їхні реєстри пакетів. Здоров'я вимірюється за тією ж версіонованою методологією, що й решта реєстру.

10 записів
З тегом «model-serving»Упорядковано за індексом здоров'я
PyPI · crates.io
99Винятковийіндекс здоров'я
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Python★ 88.2K4 серп. 2026 р.
Apache-2.04 серп. 2026 р. · метрики 2.10.0
PyPI · crates.io
98Винятковийіндекс здоров'я
basetenlabs/truss
The simplest way to serve AI/ML models in production
Python · Rust★ 1 195↓ 966.4K/міс27 серп. 2026 р.
MIT27 серп. 2026 р. · метрики 2.10.0
Go · PyPI
97Винятковийіндекс здоров'я
kserve/kserve
Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes
Go · Python★ 5 83728 серп. 2026 р.
Apache-2.028 серп. 2026 р. · метрики 2.10.0
PyPI
96Винятковийіндекс здоров'я
bentoml/BentoML
The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
Python★ 8 782↓ 256.2K/міс12 серп. 2026 р.
Apache-2.012 серп. 2026 р. · метрики 2.10.0
PyPI
95Винятковийіндекс здоров'я
mlrun/mlrun
MLRun is an open source MLOps platform for quickly building and managing continuous ML applications across their lifecycle. MLRun integrates into your development and CI/CD environment and automates the delivery of production data, ML pipelines, and online applications.
Python★ 1 6902 серп. 2026 р.
Apache-2.02 серп. 2026 р. · метрики 2.10.0
Go · npm
87Відміннийіндекс здоров'я
ome-projects/ome
Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton
Go★ 48121 лип. 2026 р.
Apache-2.021 лип. 2026 р. · метрики 2.10.0
PyPI
59Помірнийіндекс здоров'я
ai-hypercomputer/jetstream
JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs welcome).
Python★ 45118 лип. 2026 р.
Apache-2.018 лип. 2026 р. · метрики 2.10.0
PyPI · npm · Maven
59Помірнийіндекс здоров'я
hpnkv/a11
A streaming action runtime for AI agents, model serving, and multimodal APIs
C++ · Python · TypeScript★ 1↓ 17K/міс2 вер. 2026 р.
Без ліцензії2 вер. 2026 р. · метрики 2.10.0
PyPI
57Помірнийіндекс здоров'я
google/jetstream
JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs welcome).
Python★ 45118 лип. 2026 р.
Apache-2.018 лип. 2026 р. · метрики 2.10.0
PyPI
19Критичнийіндекс здоров'я
Lightning-Universe/stable-diffusion-deploy
Learn to serve Stable Diffusion models on cloud infrastructure at scale. This Lightning App shows load-balancing, orchestrating, pre-provisioning, dynamic batching, GPU-inference, micro-services working together via the Lightning Apps framework.
Python · TypeScript★ 39121 лип. 2026 р.
Apache-2.021 лип. 2026 р. · метрики 2.10.0