89 active products

NVIDIA

Publisher of DALI (NVIDIA Data Loading Library) and Tensorrt

Consolidated Product Portfolio

Nvbench

CUDA Kernel Benchmarking Library

Performance TestingOpen Source
OSMO

The developer-first platform for scaling complex Physical AI workloads across heterogeneous compute—unifying training GPUs, simulation clusters, and edge devices in a simple YAML

Mlops PlatformOpen Source
SkillSpector
SkillSpector

Security scanner for AI agent skills. Detect vulnerabilities, malicious patterns, security risks, prompt injection, data exfiltration, and supply-chain risks in Claude Code, Codex, and MCP skills before you install them.

Mcp ServerOpen Source
TensorRT-LLM

Welcome to TensorRT LLM’s Documentation! — TensorRT LLM

Llm OrchestrationOpen Source
Megatron LM
Megatron LM

Ongoing research training transformer models at scale

Llm OrchestrationOpen Source
Cosmos
Cosmos

NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

Iot PlatformOpen Source
Gontainer

Simple but powerful dependency injection container for Go projects!

Dependency ManagementOpen Source
cuDF
cuDF

cuDF - GPU DataFrame Library

Data Science NotebooksOpen Source
Megatron-Energon

Megatron's multi-modal data loader

Data IngestionOpen Source
NVSentinel

NVSentinel detects and remediates GPU faults on Kubernetes nodes

Containerization PlatformOpen Source
Container Canary

A tool for testing and validating container requirements against versioned manifests

Container ManagementOpen Source
Nvbandwidth

A tool for bandwidth measurements on NVIDIA GPUs.

Bandwidth MonitoringOpen Source
NeMo-Retriever
NeMo-Retriever

NeMo-Retriever is a Python library for scalable document content and metadata extraction.

Ai Document ExtractionOpen Source
NemoClaw
NemoClaw

Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference

Ai Agent PlatformOpen Source
TransformerEngine
TransformerEngine

Transformer Engine documentation — Transformer Engine 2.19.0

Machine Learning Model TrainingOpen Source
MinkowskiEngine
MinkowskiEngine

Minkowski Engine is an auto-diff neural network library for high-dimensional sparse tensors

Machine Learning Model TrainingOpen Source
K8s Device Plugin
K8s Device Plugin

NVIDIA device plugin for Kubernetes

Containerization PlatformOpen Source
GPU Operator
GPU Operator

NVIDIA GPU Operator creates, configures, and manages GPUs in Kubernetes

Containerization PlatformOpen Source
AIStore
AIStore

AIStore: scalable storage for AI applications

Containerization PlatformOpen Source
Nvidia Container Toolkit
Nvidia Container Toolkit

Build and run containers leveraging NVIDIA GPUs

Container ManagementOpen Source
WaveGlow
WaveGlow

A Flow-based Generative Network for Speech Synthesis

Ai Voice SynthesisOpen Source
Stable-Diffusion-WebUI-TensorRT
Stable-Diffusion-WebUI-TensorRT

TensorRT Extension for Stable Diffusion Web UI

Ai Image EditingOpen Source
Tensorrt

TensorRT is a toolkit for optimizing deep learning inference applications.

Ai GatewayCommercial
Skills
Skills

Agent Skills for NVIDIA products — install into Claude Code, Codex, and other coding agents to run Physical AI, robotics, simulation, CUDA, and RAG workflows end to end.

Ai Agent PlatformOpen Source
NeMo-Agent-Toolkit
NeMo-Agent-Toolkit

The NVIDIA NeMo Agent toolkit is an open-source library for efficiently connecting and optimizing teams of AI agents.

Ai Agent PlatformOpen Source
DALI (NVIDIA Data Loading Library)

NVIDIA DALI is a library designed for efficient data loading and preprocessing in machine learning workflows.

On Device MlCommercial
cuOpt

GPU accelerated decision optimization

Resource OptimizationOpen Source
NeMo-Speech.cpp

NeMo-Speech.cpp is a lightweight C++ inference runtime for Speech models

Model DeploymentOpen Source
NVSHMEM

NVIDIA NVSHMEM is a parallel programming interface for NVIDIA GPUs based on OpenSHMEM. NVSHMEM can significantly reduce multi-process communication and coordination overheads by allowing programmers to perform one-sided communication from within CUDA kernels and on CUDA streams.

Machine Learning Model TrainingOpen Source
GMAT

A toolkit showing GPU's all-round capability in video processing

Machine Learning Model TrainingOpen Source
Cosmos Framework

Our inference and training framework to run on the Cosmos Models

Machine Learning Model TrainingOpen Source
cuVS

cuVS - a library for vector search and clustering on the GPU

Llm OrchestrationOpen Source
FlashDreams

High-performance streaming video diffusion framework with pluggable model backends

Ai Video CreationOpen Source
SkillEvaluator

Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.

Vulnerability ManagementOpen Source
Cosmos Curator

Cosmos Curator is a powerful video curation system that processes, analyzes, and organizes video content using advanced AI models and distributed computing.

Video AnalyticsOpen Source
Sentiment Discovery

Unsupervised Language Modeling at scale for robust sentiment classification

Text Sentiment AnalysisOpen Source
Gds Nvidia Fs

NVIDIA GPUDirect Storage Driver

Storage ManagementOpen Source
Nsight Training

Training material for Nsight developer tools

Secure Code TrainingOpen Source
MDL-SDK

NVIDIA Material Definition Language SDK

Sdk GenerationOpen Source
Audio2Face-3D-SDK

High-performance C++/CUDA SDK for running Audio2Emotion and Audio2Face inference with integrated post-processing.

Sdk GenerationOpen Source
Aerial Cuda Accelerated Ran

An SDK (Software Development Kit) for building commercial-grade, AI-native, 3GPP, and O-RAN compliant 5G/6G gNB software on NVIDIA-accelerated computing platforms.

Sdk GenerationOpen Source
Soma Retargeter

SOMA BVH to humanoid robot motion retargeting library built with Newton and NVIDIA Warp

RetargetingOpen Source
Aerial Framework

A toolchain for generating high-performance, GPU-accelerated 5G/6G pipelines from Python and a modular, real-time runtime for executing the pipelines on NVIDIA Aerial™ RAN Computer platforms.

Real Time Data ProcessingOpen Source
PyProf

A GPU performance profiling tool for PyTorch models

Performance TestingOpen Source
NeMo-text-processing

NeMo text processing for ASR and TTS

Nlp Text AnalysisOpen Source
Trt LLM As Openai Windows

This reference can be used with any existing OpenAI integrated apps to run with TRT-LLM inference locally on GeForce GPU on Windows instead of cloud.

Model DeploymentOpen Source
Nvcf

Platform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.

Model DeploymentOpen Source
Kvpress

LLM KV cache compression made easy

Model DeploymentOpen Source
Kubevirt GPU Device Plugin

NVIDIA k8s device plugin for Kubevirt

K8s ManagementOpen Source
Edk2 Nvidia

NVIDIA EDK2 platform support

Iot PlatformOpen Source
Aicr

Tooling for optimized, validated, and reproducible GPU-accelerated AI runtime in Kubernetes

GitopsOpen Source
FasterTransformer
FasterTransformer

Transformer related optimization, including BERT, GPT

Generative Engine OptimizationOpen Source
Asset Harvester

NVIDIA Asset Harvester is a generative AI model for autonomous vehicle simulation that extracts reusable 3D Gaussian assets directly from captured driving scenes, even from partial or obstructed views.

Generative Engine OptimizationOpen Source
Physicsnemo Cfd

L​ibrary for using the models trained in PhysicsNeMo in Engineering and CFD workflows

Engineering AnalyticsOpen Source
Cudf Spark

NVIDIA cuDF for Apache Spark plugin - accelerate Apache Spark with GPUs

Data StreamingOpen Source
Nsight Python

Python GPU kernel profiling interface

Data ProfilingOpen Source
DCGM

NVIDIA Data Center GPU Manager (DCGM) is a project for gathering telemetry and measuring the health of NVIDIA GPUs

Data ObservabilityOpen Source
NeMo-speech-data-processor

A toolkit for processing speech data and creating speech datasets

Data CollectionOpen Source
Vgpu Device Manager

NVIDIA vGPU Device Manager manages NVIDIA vGPU devices on top of Kubernetes

Containerization PlatformOpen Source
Topograph

A toolkit for discovering cluster network topology.

Containerization PlatformOpen Source
K8s Nim Operator

An Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment.

Containerization PlatformOpen Source
Pyxis

Container plugin for Slurm Workload Manager

Container ManagementOpen Source
Libnvidia Container

NVIDIA container runtime library

Container ManagementOpen Source
Enroot

A simple yet powerful tool to turn traditional container/OS images into unprivileged sandboxes.

Container Image ManagementOpen Source
MagnumIO

Magnum IO community repo

Community BuildingOpen Source
Cloud Native Stack

Run cloud native workloads on NVIDIA GPUs

Cloud Native DevelopmentOpen Source
AI Assisted Annotation Client

Client side integration example source code and libraries for AI-Assisted Annotation SDK

Annotation LabelingOpen Source
Mellotron

Mellotron: a multispeaker voice synthesis model based on Tacotron 2 GST that can make a voice emote and sing without emotive or singing training data

Ai Voice SynthesisOpen Source
Flowtron

Flowtron is an auto-regressive flow-based generative network for text to speech synthesis with control over speech variation and style transfer

Ai Voice SynthesisOpen Source
Garak

the LLM vulnerability scanner

Ai EvaluationOpen Source
Compute Eval

Evaluating Large Language Models for CUDA Code Generation ComputeEval is a framework designed to generate and evaluate CUDA code from Large Language Models.

Ai Coding AgentOpen Source
TensorRT-Edge-LLM

High-performance, light-weight C++ LLM and VLM Inference Software for Physical AI

Llm OrchestrationOpen Source
Star-Attention

Efficient LLM Inference over Long Sequences

Llm OrchestrationOpen Source
Nim Anywhere

Accelerate your Gen AI with NVIDIA NIM and NVIDIA AI Workbench

Llm OrchestrationOpen Source
TensorRT-Model-Connect

From PyTorch model to end-to-end TensorRT inference experience in two commands—AI-native, cross-platform, and built for the best possible user experience.

Ai Agent PlatformOpen Source
OpenShell-Community

OpenShell is the safe, private runtime for autonomous AI agents.

Ai Agent PlatformOpen Source
OpenShell
OpenShell

OpenShell is the safe, private runtime for autonomous AI agents.

Ai Agent PlatformOpen Source
NeMo-Relay

Core Rust SDK for NeMo Relay observability, scope management, and runtime instrumentation.

Ai Agent PlatformOpen Source
NeMo-Agent-Toolkit-UI

The NVIDIA NeMo Agent Toolkit UI streamlines interacting with NeMo Agent Toolkit workflows in an easy-to-use web application.

Ai Agent PlatformOpen Source
Torch Harmonics

Differentiable signal processing on the sphere for PyTorch

Machine Learning Model TrainingOpen Source
pix2pixHD

Synthesizing and manipulating 2048x1024 images with conditional GANs

Machine Learning Model TrainingOpen Source
nvImageCodec

A nvImageCodec library of GPU- and CPU- accelerated codecs featuring a unified interface

Machine Learning Model TrainingOpen Source
Nvidia Resiliency Ext

NVIDIA Resiliency Package

Machine Learning Model TrainingOpen Source
Nvalchemi Toolkit

ALCHEMI Toolkit is a developer toolkit for accelerating training and inference for AI in chemistry and material science.

Machine Learning Model TrainingOpen Source
Ising-Decoding

A set of training recipes for AI Quantum Error Correction Decoders

Machine Learning Model TrainingOpen Source
DeepStream

NVIDIA DeepStream SDK Developer Guide — DeepStream documentation

Machine Learning Model TrainingOpen Source
Dataset_Utilities

NVIDIA Dataset Utilities (NVDU)

Machine Learning Model TrainingOpen Source
cuEquivariance

cuEquivariance is a math library that is a collective of low-level primitives and tensor ops to accelerate widely-used models, like DiffDock, MACE, Allegro and NEQUIP, based on equivariant neural networks. Also includes kernels for accelerated structure prediction.

Machine Learning Model TrainingOpen Source
Apex
Apex

A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch

Machine Learning Model TrainingOpen Source

Company Profile & Strategy

NVIDIA publishes 2 products at nvidia.com, including DALI (NVIDIA Data Loading Library) and Tensorrt. DALI (NVIDIA Data Loading Library): NVIDIA DALI is a library designed for efficient data loading and preprocessing in machine learning workflows.

Open-Source Footprint

87 repositories193.9K stars0 contributors

Save & manage the vendor's products in Productapp.

Save and manage this vendor's products in Productapp.

Start free trialWebsite