Alle Tags
Katalog-Tag

#moe

Alle Repositories im öffentlichen Register, die dieses Tag tragen — aus ihren GitHub-Topics oder den von ihren Paket-Registries veröffentlichten Schlagwörtern. Die Gesundheit wird nach derselben versionierten Methodik gemessen wie im übrigen Register.

11 Einträge
Getaggt als „moe“Geordnet nach Gesundheitsindex
PyPI · crates.io
99AußergewöhnlichGesundheitsindex
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Python★ 88.2K4. Aug. 2026
Apache-2.04. Aug. 2026 · Metriken 2.10.0
PyPI · crates.io
98AußergewöhnlichGesundheitsindex
sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
Python★ 31.3K↓ 144M/Monat4. Aug. 2026
Apache-2.04. Aug. 2026 · Metriken 2.10.0
PyPI
97AußergewöhnlichGesundheitsindex
flashinfer-ai/flashinfer
FlashInfer: Kernel Library for LLM Serving
Python · Cuda★ 6.263↓ 7.1M/Monat27. Aug. 2026
Apache-2.027. Aug. 2026 · Metriken 2.10.0
PyPI
94AußergewöhnlichGesundheitsindex
modelscope/ms-swift
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
Python★ 15.4K↓ 111K/Monat27. Aug. 2026
Apache-2.027. Aug. 2026 · Metriken 2.10.0
PyPI
89ExzellentGesundheitsindex
ModelCloud/GPTQModel
LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.
Python · Cuda★ 1.20717. Juli 2026
Eigene Lizenz17. Juli 2026 · Metriken 2.10.0
PyPI
88ExzellentGesundheitsindex
NVIDIA/cudnn-frontend
cuDNN Frontend is NVIDIA's modern, open-source entry point to the cuDNN library and a growing collection of high-performance open-source kernels.
Python · C++★ 88621. Juli 2026
MIT21. Juli 2026 · Metriken 2.10.0
PyPI
80ExzellentGesundheitsindex
manjunathshiva/turboquant-mlx
Extreme weight + KV cache compression for LLMs on Apple Silicon (MLX implementation of Google's TurboQuant)
Python★ 83↓ 1.116/Monat28. Aug. 2026
Eigene Lizenz28. Aug. 2026 · Metriken 2.10.0
PyPI · crates.io
63MittelGesundheitsindex
FedericoTs/quantprobe
Run a 110B on a 2016 PC with 16 GB RAM. Know your tok/s before you download. Placement beats budget: predicts speed + memory fit for any GGUF on your exact hardware, self-calibrates, emits the exact llama.cpp command — or 'quantprobe auto' does it all. Falsification-tested laws; misses published at full size. pip install quantprobe
Python★ 86↓ 6.494/Monat20. Aug. 2026
MIT20. Aug. 2026 · Metriken 2.10.0
PyPI
63MittelGesundheitsindex
RobTand/gridbook
Out-of-tree vLLM plugin and open format spec for NVFP4-CB / FP8-CB product-codebook weights — 2-6 bit-per-weight LLM quantization served on native Blackwell tensor cores.
Python · Cuda★ 10↓ 2.218/Monat15. Aug. 2026
Apache-2.015. Aug. 2026 · Metriken 2.10.0
PyPI
54MittelGesundheitsindex
jjang-ai/jangq
JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon
Python · Swift★ 21422. Juli 2026
Keine Lizenz22. Juli 2026 · Metriken 2.10.0
PyPI
50MittelGesundheitsindex
pjordanandrsn/experts4bit-qlora
QLoRA fine-tuning of fused 4-bit Mixture-of-Experts on a single small GPU (bitsandbytes Experts4bit)
Python★ 1↓ 2.064/Monat1. Aug. 2026
Eigene Lizenz1. Aug. 2026 · Metriken 2.10.0