NVIDIA
Publisher of DALI (NVIDIA Data Loading Library) and Tensorrt
Consolidated Product Portfolio
CUDA Kernel Benchmarking Library
The developer-first platform for scaling complex Physical AI workloads across heterogeneous compute—unifying training GPUs, simulation clusters, and edge devices in a simple YAML
Security scanner for AI agent skills. Detect vulnerabilities, malicious patterns, security risks, prompt injection, data exfiltration, and supply-chain risks in Claude Code, Codex, and MCP skills before you install them.
Welcome to TensorRT LLM’s Documentation! — TensorRT LLM
Ongoing research training transformer models at scale
NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.
Simple but powerful dependency injection container for Go projects!
cuDF - GPU DataFrame Library
Megatron's multi-modal data loader
NVSentinel detects and remediates GPU faults on Kubernetes nodes
A tool for testing and validating container requirements against versioned manifests
A tool for bandwidth measurements on NVIDIA GPUs.
NeMo-Retriever is a Python library for scalable document content and metadata extraction.
Run agents like Hermes, LangChain Deep Agents, and OpenClaw more securely inside NVIDIA OpenShell with managed inference
Transformer Engine documentation — Transformer Engine 2.19.0
Minkowski Engine is an auto-diff neural network library for high-dimensional sparse tensors
NVIDIA device plugin for Kubernetes
NVIDIA GPU Operator creates, configures, and manages GPUs in Kubernetes
AIStore: scalable storage for AI applications
Build and run containers leveraging NVIDIA GPUs
A Flow-based Generative Network for Speech Synthesis
TensorRT Extension for Stable Diffusion Web UI
TensorRT is a toolkit for optimizing deep learning inference applications.
Agent Skills for NVIDIA products — install into Claude Code, Codex, and other coding agents to run Physical AI, robotics, simulation, CUDA, and RAG workflows end to end.
The NVIDIA NeMo Agent toolkit is an open-source library for efficiently connecting and optimizing teams of AI agents.
NVIDIA DALI is a library designed for efficient data loading and preprocessing in machine learning workflows.
GPU accelerated decision optimization
NeMo-Speech.cpp is a lightweight C++ inference runtime for Speech models
NVIDIA NVSHMEM is a parallel programming interface for NVIDIA GPUs based on OpenSHMEM. NVSHMEM can significantly reduce multi-process communication and coordination overheads by allowing programmers to perform one-sided communication from within CUDA kernels and on CUDA streams.
A toolkit showing GPU's all-round capability in video processing
Our inference and training framework to run on the Cosmos Models
cuVS - a library for vector search and clustering on the GPU
High-performance streaming video diffusion framework with pluggable model backends
Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
Cosmos Curator is a powerful video curation system that processes, analyzes, and organizes video content using advanced AI models and distributed computing.
Unsupervised Language Modeling at scale for robust sentiment classification
NVIDIA GPUDirect Storage Driver
Training material for Nsight developer tools
NVIDIA Material Definition Language SDK
High-performance C++/CUDA SDK for running Audio2Emotion and Audio2Face inference with integrated post-processing.
An SDK (Software Development Kit) for building commercial-grade, AI-native, 3GPP, and O-RAN compliant 5G/6G gNB software on NVIDIA-accelerated computing platforms.
SOMA BVH to humanoid robot motion retargeting library built with Newton and NVIDIA Warp
A toolchain for generating high-performance, GPU-accelerated 5G/6G pipelines from Python and a modular, real-time runtime for executing the pipelines on NVIDIA Aerial™ RAN Computer platforms.
A GPU performance profiling tool for PyTorch models
NeMo text processing for ASR and TTS
This reference can be used with any existing OpenAI integrated apps to run with TRT-LLM inference locally on GeForce GPU on Windows instead of cloud.
Platform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.
LLM KV cache compression made easy
NVIDIA k8s device plugin for Kubevirt
NVIDIA EDK2 platform support
Tooling for optimized, validated, and reproducible GPU-accelerated AI runtime in Kubernetes
Transformer related optimization, including BERT, GPT
NVIDIA Asset Harvester is a generative AI model for autonomous vehicle simulation that extracts reusable 3D Gaussian assets directly from captured driving scenes, even from partial or obstructed views.
Library for using the models trained in PhysicsNeMo in Engineering and CFD workflows
NVIDIA cuDF for Apache Spark plugin - accelerate Apache Spark with GPUs
Python GPU kernel profiling interface
NVIDIA Data Center GPU Manager (DCGM) is a project for gathering telemetry and measuring the health of NVIDIA GPUs
A toolkit for processing speech data and creating speech datasets
NVIDIA vGPU Device Manager manages NVIDIA vGPU devices on top of Kubernetes
A toolkit for discovering cluster network topology.
An Operator for deployment and maintenance of NVIDIA NIMs and NeMo microservices in a Kubernetes environment.
Container plugin for Slurm Workload Manager
NVIDIA container runtime library
A simple yet powerful tool to turn traditional container/OS images into unprivileged sandboxes.
Magnum IO community repo
Run cloud native workloads on NVIDIA GPUs
Client side integration example source code and libraries for AI-Assisted Annotation SDK
Mellotron: a multispeaker voice synthesis model based on Tacotron 2 GST that can make a voice emote and sing without emotive or singing training data
Flowtron is an auto-regressive flow-based generative network for text to speech synthesis with control over speech variation and style transfer
the LLM vulnerability scanner
Evaluating Large Language Models for CUDA Code Generation ComputeEval is a framework designed to generate and evaluate CUDA code from Large Language Models.
High-performance, light-weight C++ LLM and VLM Inference Software for Physical AI
Efficient LLM Inference over Long Sequences
Accelerate your Gen AI with NVIDIA NIM and NVIDIA AI Workbench
From PyTorch model to end-to-end TensorRT inference experience in two commands—AI-native, cross-platform, and built for the best possible user experience.
OpenShell is the safe, private runtime for autonomous AI agents.
OpenShell is the safe, private runtime for autonomous AI agents.
Core Rust SDK for NeMo Relay observability, scope management, and runtime instrumentation.
The NVIDIA NeMo Agent Toolkit UI streamlines interacting with NeMo Agent Toolkit workflows in an easy-to-use web application.
Differentiable signal processing on the sphere for PyTorch
Synthesizing and manipulating 2048x1024 images with conditional GANs
A nvImageCodec library of GPU- and CPU- accelerated codecs featuring a unified interface
NVIDIA Resiliency Package
ALCHEMI Toolkit is a developer toolkit for accelerating training and inference for AI in chemistry and material science.
A set of training recipes for AI Quantum Error Correction Decoders
NVIDIA DeepStream SDK Developer Guide — DeepStream documentation
NVIDIA Dataset Utilities (NVDU)
cuEquivariance is a math library that is a collective of low-level primitives and tensor ops to accelerate widely-used models, like DiffDock, MACE, Allegro and NEQUIP, based on equivariant neural networks. Also includes kernels for accelerated structure prediction.
A PyTorch Extension: Tools for easy mixed precision and distributed training in Pytorch
Company Profile & Strategy
NVIDIA publishes 2 products at nvidia.com, including DALI (NVIDIA Data Loading Library) and Tensorrt. DALI (NVIDIA Data Loading Library): NVIDIA DALI is a library designed for efficient data loading and preprocessing in machine learning workflows.
Open-Source Footprint
Related vendors
Save & manage the vendor's products in Productapp.
Save and manage this vendor's products in Productapp.