Усі теги
Тег каталогу

#moe

Усі репозиторії публічного реєстру з цим тегом — із тем GitHub або ключових слів, які публікують їхні реєстри пакетів. Здоров'я вимірюється за тією ж версіонованою методологією, що й решта реєстру.

11 записів
З тегом «moe»Упорядковано за індексом здоров'я
PyPI · crates.io
99Винятковийіндекс здоров'я
vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Python★ 88.2K4 серп. 2026 р.
Apache-2.04 серп. 2026 р. · метрики 2.10.0
PyPI · crates.io
98Винятковийіндекс здоров'я
sgl-project/sglang
SGLang is a high-performance serving framework for large language models and multimodal models.
Python★ 31.3K↓ 144M/міс4 серп. 2026 р.
Apache-2.04 серп. 2026 р. · метрики 2.10.0
PyPI
97Винятковийіндекс здоров'я
flashinfer-ai/flashinfer
FlashInfer: Kernel Library for LLM Serving
Python · Cuda★ 6 263↓ 7.1M/міс27 серп. 2026 р.
Apache-2.027 серп. 2026 р. · метрики 2.10.0
PyPI
94Винятковийіндекс здоров'я
modelscope/ms-swift
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
Python★ 15.4K↓ 111K/міс27 серп. 2026 р.
Apache-2.027 серп. 2026 р. · метрики 2.10.0
PyPI
89Відміннийіндекс здоров'я
ModelCloud/GPTQModel
LLM model quantization (compression) toolkit with HW acceleration support for Nvidia, AMD, Intel GPU and Intel/AMD/Apple CPU via HF, vLLM, and SGLang.
Python · Cuda★ 1 20717 лип. 2026 р.
Власна ліцензія17 лип. 2026 р. · метрики 2.10.0
PyPI
88Відміннийіндекс здоров'я
NVIDIA/cudnn-frontend
cuDNN Frontend is NVIDIA's modern, open-source entry point to the cuDNN library and a growing collection of high-performance open-source kernels.
Python · C++★ 88621 лип. 2026 р.
MIT21 лип. 2026 р. · метрики 2.10.0
PyPI
80Відміннийіндекс здоров'я
manjunathshiva/turboquant-mlx
Extreme weight + KV cache compression for LLMs on Apple Silicon (MLX implementation of Google's TurboQuant)
Python★ 83↓ 1 116/міс28 серп. 2026 р.
Власна ліцензія28 серп. 2026 р. · метрики 2.10.0
PyPI · crates.io
63Помірнийіндекс здоров'я
FedericoTs/quantprobe
Run a 110B on a 2016 PC with 16 GB RAM. Know your tok/s before you download. Placement beats budget: predicts speed + memory fit for any GGUF on your exact hardware, self-calibrates, emits the exact llama.cpp command — or 'quantprobe auto' does it all. Falsification-tested laws; misses published at full size. pip install quantprobe
Python★ 86↓ 6 494/міс20 серп. 2026 р.
MIT20 серп. 2026 р. · метрики 2.10.0
PyPI
63Помірнийіндекс здоров'я
RobTand/gridbook
Out-of-tree vLLM plugin and open format spec for NVFP4-CB / FP8-CB product-codebook weights — 2-6 bit-per-weight LLM quantization served on native Blackwell tensor cores.
Python · Cuda★ 10↓ 2 218/міс15 серп. 2026 р.
Apache-2.015 серп. 2026 р. · метрики 2.10.0
PyPI
54Помірнийіндекс здоров'я
jjang-ai/jangq
JANG — GGUF for MLX. YOU MUST USE JANG_Q RUNTIME. Adaptive Mixed-Precision Quantization + Runtime for Apple Silicon
Python · Swift★ 21422 лип. 2026 р.
Без ліцензії22 лип. 2026 р. · метрики 2.10.0
PyPI
50Помірнийіндекс здоров'я
pjordanandrsn/experts4bit-qlora
QLoRA fine-tuning of fused 4-bit Mixture-of-Experts on a single small GPU (bitsandbytes Experts4bit)
Python★ 1↓ 2 064/міс1 серп. 2026 р.
Власна ліцензія1 серп. 2026 р. · метрики 2.10.0