NVIDIA technologies
Independent profiles of NVIDIA products, platforms, libraries and programs: what each one is, what it needs, what it is not, and its official sources.
- Software suite
NVIDIA AI Enterprise
NVIDIA AI Enterprise is a commercial, supported software suite that bundles NVIDIA's AI frameworks, NIM microservices and SDKs with GPU drivers, Kubernetes operators and Run:ai orchestration, adding release branches, security patching and SLA-backed support for production AI.
- Inference & runtime software
- Operations & orchestration
- Models & frameworks
- Product
NVIDIA AI Workbench
AI Workbench is NVIDIA's free tool for running containerized, Git-managed AI projects on a laptop, a GPU workstation, a remote server or a cloud instance with the same interface. It handles containers, GPU drivers and remote connections; the latest release, 2026.06.5, shipped in July 2026.
- Applications & solutions
- Operations & orchestration
- Accelerated computing
- Platform
NVIDIA BioNeMo
BioNeMo is NVIDIA's development platform for AI in biology and drug discovery: open models, training recipes, libraries, NIM microservices and, since June 2026, an agent toolkit. The older BioNeMo Framework container is archived; NVIDIA now points users to BioNeMo Recipes on GitHub.
- Applications & solutions
- Models & frameworks
- Inference & runtime software
- Hardware family
NVIDIA BlueField
NVIDIA BlueField is a family of data processing units (DPUs) that sit in servers and storage systems and run networking, storage and security services on their own processors and accelerators, so host CPUs and GPUs are left for application work.
- Networking, power & facilities
- Accelerated computing
- Reference design
NVIDIA Blueprints
NVIDIA Blueprints are open reference workflows for agentic and generative AI use cases such as retrieval-augmented generation, video search and summarization, and research agents. Each provides sample code, deployment files and documentation built on NIM microservices and other NVIDIA libraries, meant to be adapted.
- Applications & solutions
- Platform
NVIDIA Cosmos
NVIDIA Cosmos is an open platform of world foundation models and tools for physical AI. Cosmos 3 reasons over images and video and generates video, sound and robot actions; Curator, Evaluator and Cosmos Framework cover data, scoring and post-training. Licenses differ by model (OpenMDW 1.1 for Cosmos 3).
- Models & frameworks
- Inference & runtime software
- Software suite
NVIDIA CUDA Toolkit
The NVIDIA CUDA Toolkit is the development kit for programming NVIDIA GPUs: a compiler, runtime and driver APIs, math and parallel libraries, debugging and profiling tools, and documentation. The current release is CUDA 13.4 Update 1, and the GPU driver is now installed separately.
- Accelerated computing
- Library
NVIDIA CUDA-X Data Science
NVIDIA CUDA-X Data Science, known until August 2026 as RAPIDS, is a collection of open source GPU libraries for data science: cuDF for dataframes, cuML for machine learning and cuGraph for graph analytics. Several of them can speed up existing pandas, scikit-learn or NetworkX code without code changes.
- Accelerated computing
- Applications & solutions
- Library
NVIDIA cuOpt
NVIDIA cuOpt is an open source, GPU-accelerated optimization engine for routing, linear programming and quadratic programming, with mixed integer and conic programming in beta. It runs as a Python or C library or as a self-hosted server, and plugs into modeling tools such as AMPL, CVXPY, PuLP, GAMSPy and JuMP.
- Applications & solutions
- Accelerated computing
- Framework
NVIDIA DeepStream SDK
DeepStream is NVIDIA's toolkit for building real-time video and multi-sensor analytics pipelines on GPUs and Jetson devices. It is based on GStreamer and part of Metropolis. Its source code has been on GitHub under Apache-2.0 since version 9.0, while the prebuilt runtime libraries stay under an NVIDIA license.
- Applications & solutions
- Inference & runtime software
- Accelerated computing
- Platform
NVIDIA DGX
NVIDIA DGX is NVIDIA's own line of AI systems, from the DGX Spark desktop to rack-scale DGX SuperPOD clusters, delivered together with NVIDIA operations software, reference architectures and support, so an organization can build AI infrastructure on one validated stack.
- Accelerated computing
- Operations & orchestration
- Platform
NVIDIA DSX
NVIDIA DSX is a platform, not a single product: a set of reference designs, simulation tools, open-source operations software, power management and data-exchange schemas for designing, building and running AI factories, with each part usable by different partners in the build.
- Networking, power & facilities
- Operations & orchestration
- Applications & solutions
- Programs & resources
- Framework
NVIDIA Dynamo
NVIDIA Dynamo is an open source framework that coordinates generative AI inference across many GPUs and nodes. It sits above engines such as vLLM, SGLang and TensorRT LLM and adds disaggregated serving, KV-cache-aware routing, cache offload and latency-driven autoscaling.
- Inference & runtime software
- Operations & orchestration
- Product
NVIDIA Dynamo-Triton
NVIDIA Dynamo-Triton, formerly Triton Inference Server, is open source inference serving software that runs models from many frameworks, including TensorRT, PyTorch, ONNX, OpenVINO, Python and RAPIDS FIL, on GPUs and CPUs behind HTTP/REST and gRPC APIs.
- Inference & runtime software
- Platform
NVIDIA Holoscan SDK
Holoscan is NVIDIA's open source SDK for real-time AI processing of sensor streams, such as surgical video, ultrasound, cameras and radio signals, at the edge or in the cloud. It runs on IGX, Jetson, DGX Spark and x86 systems with NVIDIA GPUs and is licensed under Apache-2.0.
- Applications & solutions
- Inference & runtime software
- Accelerated computing
- Program
NVIDIA Inception
NVIDIA Inception is a free program for AI startups. Members get training, developer tools, preferred pricing on select NVIDIA hardware and software, partner offers, cloud credits, investor exposure based on eligibility, and marketing support. It charges no fees, takes no equity and has no deadlines or cohorts.
- Programs & resources
- Platform
NVIDIA Isaac
NVIDIA Isaac is an open robotics development platform: simulation and robot learning frameworks (Isaac Sim, Isaac Lab), CUDA-accelerated ROS 2 packages (Isaac ROS), robot foundation models (Isaac GR00T) and reference workflows for building mobile robots, robot arms and humanoids.
- Applications & solutions
- Models & frameworks
- Inference & runtime software
- Hardware family
NVIDIA Jetson
NVIDIA Jetson is a family of compact computer modules and developer kits with NVIDIA GPUs for running AI inside robots, drones, cameras and other edge devices. It spans entry modules up to the Blackwell-based Jetson Thor series and runs the JetPack software stack.
- Accelerated computing
- Inference & runtime software
- Platform
NVIDIA Metropolis
NVIDIA Metropolis is a vision AI application platform and partner ecosystem. It bundles models, libraries and blueprints, such as the Video Search and Summarization (VSS) blueprint, DeepStream and TAO, for building video analytics agents that turn camera streams into events, alerts, search and reports.
- Applications & solutions
- Models & frameworks
- Inference & runtime software
- Platform
NVIDIA Mission Control
NVIDIA Mission Control is operations software for AI factories built on DGX and GB200/GB300 NVL72 systems. It brings cluster provisioning, Slurm and Kubernetes scheduling, health checks, automated recovery, power policies and building management integration into one supported control plane.
- Operations & orchestration
- Software suite
NVIDIA NeMo
NVIDIA NeMo is an open suite of libraries for preparing data, training and post-training models, evaluating them and adding guardrails to AI agents. Its containerized NeMo Microservices reached their announced sunset date of October 1, 2026; NVIDIA's docs still list the open source NeMo Framework.
- Models & frameworks
- Operations & orchestration
- Model family
NVIDIA Nemotron
NVIDIA Nemotron is NVIDIA's family of open AI models for building agents: reasoning models in several sizes plus models for vision, retrieval, speech and safety. NVIDIA publishes the weights, much of the training data and the training recipes, and the models run on common open inference engines or as NIM.
- Models & frameworks
- Microservice
NVIDIA NIM
NVIDIA NIM packages an AI model, an inference engine and its runtime into a container with standard APIs, so teams can self-host models on NVIDIA GPUs in the cloud, a data center, a workstation or at the edge instead of building their own serving stack.
- Inference & runtime software
- Operations & orchestration
- Software suite
NVIDIA Nsight Developer Tools
NVIDIA Nsight is NVIDIA's family of developer tools for profiling, debugging and analyzing software on NVIDIA GPUs. Nsight Systems shows a timeline of CPU, GPU and system activity, Nsight Compute profiles individual CUDA kernels, and Nsight Graphics covers graphics, alongside debuggers and IDE plugins.
- Operations & orchestration
- Accelerated computing
- Platform
NVIDIA Omniverse
NVIDIA Omniverse is a set of GPU-accelerated libraries, APIs and services for building physically based 3D simulations and digital twins on OpenUSD data. NVIDIA states it has been free for development, production and redistribution since May 2026; enterprise support needs an AI Enterprise license.
- Applications & solutions
- Inference & runtime software
- Accelerated computing
- Platform Preview
NVIDIA OpenShell
NVIDIA OpenShell is an open source runtime that runs AI agents inside isolated sandboxes and enforces declared policies on their file access, system calls, network connections, credentials and model calls. It is early software: NVIDIA called it an early preview in May 2026, and the docs are at version 0.1.2.
- Operations & orchestration
- Applications & solutions
- Platform
NVIDIA RTX Remix
RTX Remix is NVIDIA's open source modding platform for remastering classic DirectX 8 and 9 games with path tracing, DLSS and new physically based assets. It left beta in March 2025; version 1.5 shipped in June 2026. Its parts carry different open source licenses.
- Applications & solutions
- Accelerated computing
- Platform
NVIDIA Run:ai
NVIDIA Run:ai is a Kubernetes-based platform that pools GPUs and schedules AI workloads across teams using quotas, priorities and fair sharing, so a shared cluster can serve notebooks, training and inference without each team owning fixed hardware.
- Operations & orchestration
- Platform
NVIDIA Spectrum-X Ethernet
NVIDIA Spectrum-X Ethernet is a networking platform for AI clusters that pairs Spectrum-X switches with SuperNICs in the GPU servers, so congestion control, adaptive routing and telemetry work end to end on standards-based Ethernet.
- Networking, power & facilities
- Library
NVIDIA TensorRT LLM
NVIDIA TensorRT LLM (often written TensorRT-LLM) is an open source library that speeds up large language model and visual generation inference on NVIDIA GPUs, using custom kernels, quantization, in-flight batching, paged KV cache, speculative decoding and multi-GPU parallelism behind a Python LLM API.
- Inference & runtime software
No technology matches. Clear the filter, or try the search above for a business problem.
Thank you. Your correction was sent.
The editors check it against the sources. If you left an email address, they may reply about it.