PyPI · crates.io99Exceptionalhealth index
Rust · Python · Go★ 7,886↓ 59.4K/moAug 28, 2026
PyPI · npm98Exceptionalhealth index

NVIDIA-NeMo/GuardrailsNeMo Guardrails is an open-source toolkit for easily adding programmable guardrails to LLM-based conversational systems.
Python★ 6,930↓ 394.8K/moAug 12, 2026
PyPI98Exceptionalhealth index
Python★ 17.4K↓ 234.4K/moAug 12, 2026
PyPI98Exceptionalhealth index

NVIDIA/cumlNVIDIA cuML: GPU-Accelerated Machine Learning
Python · C++ · Cuda★ 5,251Aug 12, 2026
PyPI98Exceptionalhealth index

NVIDIA/warpA Python framework for GPU-accelerated simulation, robotics, and machine learning.
Python · C++★ 7,037↓ 1.2M/moAug 28, 2026
npm97Exceptionalhealth index

NVIDIA/NemoClawRun agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference
TypeScript★ 22.1K↓ 264/moAug 5, 2026
PyPI97Exceptionalhealth index
Python · Cuda★ 6,263↓ 7.1M/moAug 27, 2026
PyPI96Exceptionalhealth index
C++ · Cuda★ 2,420Jul 15, 2026
PyPI · crates.io95Exceptionalhealth index
Python · Rust★ 374Jul 26, 2026
crates.io95Exceptionalhealth index
ai-dynamo/modelexpressModel Express is a Rust-based component meant to be placed next to existing model inference systems to speed up their startup times and improve overall performance.
Python · Rust★ 103↓ 246.6K/moAug 1, 2026
PyPI · crates.io94Exceptionalhealth index

NVIDIA/OpenShellOpenShell is the safe, private runtime for autonomous AI agents.
Rust★ 8,323↓ 63.7K/moAug 22, 2026
PyPI94Exceptionalhealth index

NVIDIA/cutlassCUDA Templates and Python DSLs for High-Performance Linear Algebra
C++ · Cuda · Python★ 10.3K↓ 10.2K/moAug 28, 2026
PyPI · npm93Exceptionalhealth index
Python · MDX★ 1,055↓ 406.4K/moJul 18, 2026
PyPI · crates.io · npm93Exceptionalhealth index
NVIDIA/NeMo-RelayMulti-language agent runtime and library for execution scope management, lifecycle events, and middleware on tool and LLM calls.
Rust★ 87↓ 626.3K/moAug 1, 2026
PyPI93Exceptionalhealth index
Python · Jupyter Notebook · C++★ 2,984Jul 30, 2026
FluidInference/FluidAudioFrontier CoreML audio models in your apps — text-to-speech, speech-to-text, voice activity detection, and speaker diarization. In Swift, powered by SOTA open source.
Swift★ 2,467Jul 16, 2026
Go92Excellenthealth index

leptonai/gpudGPUd automates monitoring, diagnostics, and issue identification for GPUs
Go★ 492Sep 5, 2026
PyPI91Excellenthealth index
Python · Shell★ 33Jul 24, 2026
PyPI91Excellenthealth index
Python★ 523↓ 10.6K/moJul 29, 2026
Go91Excellenthealth index
Go★ 1,510Jul 22, 2026
NVIDIA/nvcfPlatform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.
Go · Rust★ 180Jul 17, 2026
PyPI90Excellenthealth index

rbonghi/jetson_stats📊 Simple package for monitoring and control your NVIDIA Jetson [Orin, Xavier, Nano, TX] series
Python★ 2,608↓ 39.4K/moAug 9, 2026
Go89Excellenthealth index

defilantech/LLMKubeKubernetes operator for self-hosted LLM inference across a heterogeneous GPU fleet: NVIDIA CUDA, AMD Vulkan, and Apple Silicon Metal. Runtimes: llama.cpp, vLLM, TGI, mlx-server. Multi-GPU sharding, model caching, OpenAI-compatible endpoints. Apache-2.0, run across homelab and on-prem fleets, actively developed.
Go★ 207Sep 5, 2026
PyPI88Excellenthealth index

NVIDIA/TensorRTNVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.
C++★ 13.3K↓ 1M/moAug 27, 2026
PyPI88Excellenthealth index
NVIDIA/cudnn-frontendcuDNN Frontend is NVIDIA's modern, open-source entry point to the cuDNN library and a growing collection of high-performance open-source kernels.
Python · C++★ 886Jul 21, 2026
Go86Excellenthealth index
NexusGPU/tensor-fusionTensor Fusion is a state-of-the-art GPU virtualization and pooling solution designed to optimize GPU cluster utilization to its fullest potential.
Go★ 158Jul 17, 2026
Go · npm86Excellenthealth index
TypeScript · Go · MDX★ 542Jul 28, 2026
PyPI84Excellenthealth index

NVIDIA/ncclOptimized primitives for collective multi-GPU communication
C++ · Cuda · C★ 4,988↓ 634.8K/moAug 12, 2026
Go84Excellenthealth index
nexusgpu/tensor-fusion-operatorTensor Fusion is a state-of-the-art GPU virtualization and pooling solution designed to optimize GPU cluster utilization to its fullest potential.
Go★ 158Jul 15, 2026
npm84Excellenthealth index
HTML · JavaScript★ 2,186↓ 5,344/moJul 23, 2026