PyPI92Excellenthealth index
C++ · C★ 377Aug 8, 2026
Julia · JavaScript★ 2,791Jul 25, 2026
npm91Excellenthealth index

catboost/catboostA fast, scalable, high performance Gradient Boosting on Decision Trees library, used for ranking, classification, regression and other machine learning tasks for Python, R, Java, C++. Supports computation on CPU and GPU.
C++ · Python★ 9,064↓ 7,903/moAug 12, 2026
PyPI91Excellenthealth index
Python · Shell★ 33Jul 24, 2026
PyPI91Excellenthealth index
opesci/devitoDSL and compiler framework for automated finite-differences and stencil computation
Python★ 708Jul 17, 2026
crates.io91Excellenthealth index
snipsco/tractTiny, no-nonsense, self-contained, Tensorflow and ONNX inference
Rust★ 2,991↓ 912.8K/moJul 16, 2026
PyPI91Excellenthealth index
Python★ 523↓ 10.6K/moJul 29, 2026
NVIDIA/nvcfPlatform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.
Go · Rust★ 180Jul 17, 2026
PyPI90Excellenthealth index

plasma-umass/scaleneScalene: a high-performance, high-precision CPU, GPU, and memory profiler for Python with AI-powered optimization proposals
Python · JavaScript★ 13.5K↓ 390.1K/moAug 12, 2026
PyPI90Excellenthealth index

rbonghi/jetson_stats📊 Simple package for monitoring and control your NVIDIA Jetson [Orin, Xavier, Nano, TX] series
Python★ 2,608↓ 39.4K/moAug 9, 2026
PyPI90Excellenthealth index
Python · Cuda★ 386Jul 20, 2026

CliMA/ClimaLand.jlModular, GPU-capable land surface model of the CliMA Earth System Model, designed for data-driven parameterizations
Julia★ 73Aug 17, 2026
PyPI89Excellenthealth index
Python★ 2,121Jul 21, 2026
PyPI89Excellenthealth index
Qiskit/qiskit-aerAer is a high performance simulator for quantum circuits that includes noise models
C++ · Python★ 681Jul 30, 2026
Go89Excellenthealth index

defilantech/LLMKubeKubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 207Sep 5, 2026
PyPI89Excellenthealth index
Jupyter Notebook★ 28.1KAug 5, 2026
Go89Excellenthealth index
gogpu/gogpuPure Go GPU Application Framework — windowing, input, lifecycle, platform abstraction. Part of the GoGPU ecosystem.
Go★ 341Jul 17, 2026
Go89Excellenthealth index
Go★ 157Jul 17, 2026
PyPI89Excellenthealth index
Python★ 60Jul 27, 2026
PyPI88Excellenthealth index

Andyyyy64/whichllmFind the local LLM that actually runs and performs best on your hardware. Ranked by real, recency-aware benchmarks, not parameter count. One command, run it instantly.
Python★ 6,463Aug 24, 2026
PyPI88Excellenthealth index
NVIDIA/cudnn-frontendcuDNN Frontend is NVIDIA's modern, open-source entry point to the cuDNN library and a growing collection of high-performance open-source kernels.
Python · C++★ 886Jul 21, 2026
crates.io · PyPI · npm87Excellenthealth index

AlexsJones/llmfitHundreds of models & providers. One command to find what runs on your hardware.
Rust★ 31.1K↓ 2,251/moAug 5, 2026
PyPI87Excellenthealth index
arbor-sim/arborThe Arbor multi-compartment neural network simulation library.
C++ · AGS Script★ 136Jul 25, 2026
PyPI · Maven · npm +187Excellenthealth index

h2oai/h2o-3H2O is an Open Source, Distributed, Fast & Scalable Machine Learning Platform: Deep Learning, Gradient Boosting (GBM) & XGBoost, Random Forest, Generalized Linear Modeling (GLM with Elastic Net), K-Means, PCA, Generalized Additive Models (GAM), RuleFit, Support Vector Machine (SVM), Stacked Ensembles, Automatic Machine Learning (AutoML), etc.
Jupyter Notebook · HTML · Java★ 7,495↓ 3,086/moAug 24, 2026
npm87Excellenthealth index
TypeScript · Swift · Kotlin★ 9,582↓ 2.7M/moAug 27, 2026
crates.io87Excellenthealth index
paiml/aprenderNext Generation Machine Learning, Statistics and Deep Learning in PURE Rust
Rust · HTML★ 110↓ 3,595/moJul 28, 2026
npm87Excellenthealth index

withcatai/node-llama-cppRun AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level
TypeScript★ 2,162↓ 3.6M/moAug 27, 2026
npm87Excellenthealth index

zeux/meshoptimizerMesh optimization library that makes meshes smaller and faster to render
C++ · JavaScript★ 8,209↓ 31.8M/moAug 12, 2026
Go86Excellenthealth index
NexusGPU/tensor-fusionTensor Fusion is a state-of-the-art GPU virtualization and pooling solution designed to optimize GPU cluster utilization to its fullest potential.
Go★ 158Jul 17, 2026
PyPI86Excellenthealth index
C++★ 236Jul 21, 2026