All tags
Catalogue tag

#inference

Every repository in the public record carrying this tag — from its GitHub topics or the keywords its package registries publish. Health is measured under the same versioned methodology as the rest of the record.

106 records
Tagged “inference”Ranked by health index
Go
56Moderatehealth index
wentbackward/llm-proxy
A superfast proxy and smart load-balancer for AI Inference — virtualize models, share local and provider back-ends, optimal caching, lock-in sampling parameters, debug message flow and obtain OTel metrics
Go★ 8Jul 17, 2026
MITJul 17, 2026 · metrics 2.10.0
PyPI
53Moderatehealth index
h2non/filetype.py
Small, dependency-free, fast Python package to infer binary file types checking the magic numbers signature
Python★ 769Jul 17, 2026
MITJul 17, 2026 · metrics 2.10.0
PyPI
51Moderatehealth index
ToPo-ToPo-ToPo/local-llm-server
ローカルLLM(mlx / mlx-vlm / llama.cpp / router)を OpenAI 互換 API として起動・管理する軽量サーバー
Python★ 0↓ 4,302/moJul 19, 2026
Apache-2.0Jul 19, 2026 · metrics 2.10.0
NuGet
50Moderatehealth index
mdesalvo/OWLSharp
Lightweight and friendly .NET library for realizing modern Semantic Web applications (OWL2, SWRL)
C#★ 16Jul 22, 2026
Apache-2.0Jul 22, 2026 · metrics 2.10.0
Go
50Moderatehealth index
thalesfsp/inference
Provides building blocks to integrate with AI / LLM providers and built-in, common, providers.
Go★ 0Jul 17, 2026
MITJul 17, 2026 · metrics 2.10.0
npm
47Weakhealth index
stfurkan/bitgpu
Fast WebGPU runtime for 1-bit (binary-weight) LLMs in the browser. Bit-exact, zero runtime dependencies.
TypeScript · WGSL · Python★ 19↓ 3,349/moJul 31, 2026
MITJul 31, 2026 · metrics 2.10.0
PyPI
45Weakhealth index
mihir0209/AI_engine
Free AI inference router for developers - 21 providers, OpenAI-compatible
Python★ 4↓ 10.8K/moJul 14, 2026
MITJul 14, 2026 · metrics 2.10.0
PyPI
42Weakhealth index
llaa33219/outo-llms
Deploy local LLMs behind your own managed API server - vLLM or llama.cpp in isolated environments, API keys, workspaces, usage tracking
Python · JavaScript · CSS★ 0↓ 2,686/moJul 24, 2026
Apache-2.0Jul 24, 2026 · metrics 2.10.0
Packagist
34At Riskhealth index
JetBrains/phpstorm-stubs
PHP runtime & extensions header files for PhpStorm
PHP★ 1,394↓ 2.2M/moAug 27, 2026
Apache-2.0Aug 27, 2026 · metrics 2.10.0
npm · crates.io
34At Riskhealth index
compose-market/sdk
No repository description published.
TypeScript★ 0↓ 1,662/moSep 5, 2026
MITSep 5, 2026 · metrics 2.10.0
PyPI
34At Riskhealth index
openvinotoolkit/openvino
OpenVINO™ is an open source toolkit for optimizing and deploying AI inference
C++★ 10.6KAug 12, 2026
Apache-2.0Aug 12, 2026 · metrics 2.10.0
PyPI
28At Riskhealth index
Saganaki22/MisoTTS-ComfyUI
Miso TTS 8B nodes for ComfyUI - Sesame-style CSM text-to-speech with Mimi audio tokens, optional reference-audio continuation, Whisper transcription, and Aimdo/VRAM-management integration.
Python★ 7Aug 4, 2026
MITAug 4, 2026 · metrics 2.10.0
PyPI
21At Riskhealth index
declare-lab/reccon
This repository contains the dataset and the PyTorch implementations of the models from the paper Recognizing Emotion Cause in Conversations.
Python★ 190Jul 15, 2026
No licenseJul 15, 2026 · metrics 2.10.0
PyPI
21At Riskhealth index
nactttch/g2n
A torch.compile backend that pays attention: custom FX fusion passes, a Triton LayerNorm kernel, and a persistent compile cache. Apache-2.0 free core of the g2n platform.
Python★ 0↓ 2,377/moAug 1, 2026
Apache-2.0Aug 1, 2026 · metrics 2.10.0
PyPI · Maven
19Criticalhealth index
awslabs/multi-model-server
Multi Model Server is a tool for serving neural net models for inference
Java · Python★ 1,024↓ 352.2K/moAug 13, 2026
Apache-2.0Aug 13, 2026 · metrics 2.10.0