PyPI88Excellenthealth index
huggingface/transformers🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
Python★ 162.7KJul 20, 2026
crates.io · PyPI88Excellenthealth index
vllm-project/vllmA high-throughput and memory-efficient inference and serving engine for LLMs
Python★ 86.1K↓ 0/moJul 13, 2026
PyPI86Excellenthealth index
tenstorrent/tt-metal:metal: TT-NN operator library, and TT-Metalium low level kernel programming model.
C++ · Python★ 1,584Jul 17, 2026
Go · npm85Excellenthealth index
mudler/LocalAILocalAI is the open-source AI engine. Run any model - LLMs, vision, voice, image, video - on any hardware. No GPU required.
Go · JavaScript★ 47.6KJul 15, 2026
PyPI85Excellenthealth index
strands-agents/toolsA set of tools that gives agents powerful capabilities.
Python★ 1,132Jul 21, 2026
Maven · npm84Goodhealth index
langchain4j/langchain4jLangChain4j is an idiomatic, open-source Java library for building LLM-powered applications on the JVM. It offers a unified API over popular LLM providers and vector stores, and makes implementing tool calling (including MCP support), agents and RAG easy. It integrates seamlessly with enterprise Java frameworks like Quarkus and Spring Boot.
Java★ 12.6KJul 20, 2026
modelscope/ms-swiftUse PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
Python★ 14.8K↓ 103.3K/moJul 15, 2026
pytorch/aoPyTorch native quantization and sparsity for training and inference
Python · C++★ 2,909Jul 21, 2026
modelscope/swiftUse PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, Phi4, ...) (AAAI 2025).
Python★ 14.8K↓ 0/moJul 14, 2026
crates.io · PyPI82Goodhealth index
sgl-project/sglangSGLang is a high-performance serving framework for large language models and multimodal models.
Python★ 30.3K↓ 274M/moJul 14, 2026
PyPI · npm82Goodhealth index
unslothai/unslothUnsloth is a local UI for training and running Gemma 4, Qwen3.6, DeepSeek, Kimi, GLM and other models.
Python · TypeScript★ 68.5K↓ 2.3M/moJul 20, 2026
PyPI · npm81Goodhealth index
astrbotdevs/astrbotAI Agent Assistant & development framework that integrates lots of IM platforms, LLMs, plugins and AI feature, and can be your openclaw alternative. ✨
Python · Vue★ 36.5K↓ 25.1K/moJul 17, 2026
ollama/ollamaGet up and running with Kimi-K2.6, GLM-5.1, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Go · C★ 176.2KJul 15, 2026
jmorganca/ollamaGet up and running with Kimi-K2.6, GLM-5.1, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Go · C★ 176.3KJul 17, 2026
yamadashy/repomix📦 Repomix is a powerful tool that packs your entire repository into a single, AI-friendly file. Perfect for when you need to feed your codebase to Large Language Models (LLMs) or other AI tools like Claude, ChatGPT, DeepSeek, Perplexity, Gemini, Gemma, Llama, Grok, and more.
TypeScript★ 27.2K↓ 310.9K/moJul 17, 2026
lemonade-sdk/lemonadeLemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk
C++ · Python · TypeScript★ 4,964Jul 17, 2026
SciSharp/LLamaSharpA C#/.NET library to run LLM (🦙LLaMA/LLaVA) on your local device efficiently.
C# · JavaScript★ 3,742Jul 18, 2026
Go · npm73Goodhealth index
ome-projects/omeOpen Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton
Go★ 481Jul 21, 2026
hexastack/hexabotHexabot v3 is an AI workflow automation platform, combining workflows, actions, agents, and conversational channels in one runtime.
TypeScript★ 1,091Jul 17, 2026
NuGet69Moderatehealth index
awaescher/OllamaSharpThe easiest way to use Ollama in .NET
C#★ 1,391Jul 15, 2026
npm69Moderatehealth index
llm-exe/llm-exeA package that provides simplified base components to make building and maintaining LLM-powered applications easier.
TypeScript★ 133↓ 2,631/moJul 15, 2026
npm · Maven69Moderatehealth index
mybigday/llama.rnReact Native binding of llama.cpp
C++ · C★ 1,000↓ 57.2K/moJul 17, 2026
Packagist68Moderatehealth index
symfony/ai-platformPHP library for interacting with AI platform provider.
PHP★ 52↓ 197.4K/moJul 15, 2026
hybridgroup/yzmaGo with your own intelligence - Go applications that directly integrate llama.cpp for local inference using hardware acceleration.
Go★ 520Jul 17, 2026
PyPI63Moderatehealth index
Aider-AI/aideraider is AI pair programming in your terminal
Python★ 47.6K↓ 848.4K/moJul 21, 2026
invergent-ai/surogateTraining/Fine-tuning at the speed of light
C++ · Python · Cuda★ 806↓ 0/moJul 21, 2026
PyPI60Moderatehealth index
jagmarques/nexusquantTraining-free KV cache compression via E8 lattice VQ. 2-bit KV that preserves retrieval (30/30 NIAH vs TurboQuant 0/30). Calibration-free, 9 architectures validated.
Python★ 25Jul 16, 2026
RubyGems · npm60Moderatehealth index
yoshoku/llama_cpp.rbllama_cpp.rb provides Ruby bindings for llama.cpp
C★ 235Jul 15, 2026
npm59Moderatehealth index
ggml-org/llama.vscodeVS Code extension for LLM-assisted code/text completion
TypeScript★ 1,445Jul 17, 2026
npm59Moderatehealth index
therealtimex/node-llama-cppRun AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 1↓ 41.2K/moJul 17, 2026