全部标签
目录标签

#model-serving

公开记录中带有此标签的全部仓库——标签来自其 GitHub 主题或软件包注册表发布的关键词。健康度量遵循与记录其余部分相同的版本化方法论。

7 条记录
标签为“model-serving”按健康指数排序
crates.io · PyPI
88优秀健康指数
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Python★ 86.1K↓ 0/月2026年7月13日
Apache-2.02026年7月13日 · 指标 1.13.0
Go · PyPI
87优秀健康指数
kubeflow/kfserving
Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes
Go · Python★ 5,7052026年7月17日
Apache-2.02026年7月17日 · 指标 1.13.0
Go · PyPI
85优秀健康指数
kserve/kserve
Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes
Go · Python★ 5,7052026年7月17日
Apache-2.02026年7月17日 · 指标 1.13.0
Go · npm
73良好健康指数
ome-projects/ome
Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton
Go★ 4812026年7月21日
Apache-2.02026年7月21日 · 指标 1.13.0
PyPI
55中等健康指数
ai-hypercomputer/jetstream
JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs welcome).
Python★ 4512026年7月18日
Apache-2.02026年7月18日 · 指标 1.13.0
PyPI
55中等健康指数
google/jetstream
JetStream is a throughput and memory optimized engine for LLM inference on XLA devices, starting with TPUs (and GPUs in future -- PRs welcome).
Python★ 4512026年7月18日
Apache-2.02026年7月18日 · 指标 1.13.0
PyPI
22危急健康指数
Lightning-Universe/stable-diffusion-deploy
Learn to serve Stable Diffusion models on cloud infrastructure at scale. This Lightning App shows load-balancing, orchestrating, pre-provisioning, dynamic batching, GPU-inference, micro-services working together via the Lightning Apps framework.
Python · TypeScript★ 3912026年7月21日
Apache-2.02026年7月21日 · 指标 1.13.0