NVIDIA DeepStream
NVIDIA DeepStream is a GPU-accelerated software development kit for building real-time streaming analytics pipelines, principally for video.
Explore NVIDIA through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of NVIDIA.
Showing 61-120 of 134 articles
NVIDIA DeepStream is a GPU-accelerated software development kit for building real-time streaming analytics pipelines, principally for video.
NVIDIA Dynamo is an open-source, low-latency distributed inference serving framework designed to deploy and scale generative AI and reasoning models across large GPU clusters.
NVIDIA Exemplar Cloud is a validation program run by NVIDIA that certifies cloud providers whose GPU clusters reproduce at least 95% of the training throughput NVIDIA measures on its own reference architecture…
NVIDIA Feynman is the data center GPU architecture that NVIDIA has placed on its public roadmap as the successor to Rubin and Rubin Ultra, with a target of around 2028.
The NVIDIA GB200 NVL72 is a rack-scale AI computing system that packages 72 Blackwell GPUs and 36 Grace CPUs into a single liquid-cooled rack
The NVIDIA GB300 NVL72 is a liquid-cooled, rack-scale AI computing system that integrates 72 NVIDIA Blackwell Ultra (B300) GPUs and 36 Arm-based NVIDIA Grace CPUs into a single NVLink fabric, delivering 1.1…
The NVIDIA GH200 Grace Hopper Superchip is a single-module processor from NVIDIA that combines a 72-core Grace Arm CPU with a Hopper-generation H100-class GPU on one package
NVIDIA Grace is an Arm-based data-center central processing unit (CPU) from Nvidia, built from 72 Arm Neoverse V2 cores with co-packaged LPDDR5X memory and a coherent NVLink-C2C interconnect that links it to…
NVIDIA Groq 3 LPX is a rack-scale inference accelerator that NVIDIA introduced at GTC 2026, built around 256 Groq Language Processing Units and designed to sit beside Vera Rubin NVL72 racks as a dedicated…
NVIDIA H100 (also called the H100 Tensor Core GPU) is a data-center graphics processing unit built by NVIDIA on the Hopper microarchitecture, fabricated with over 80 billion transistors on a custom TSMC 4N (4…
The NVIDIA H20 is a data-center graphics processing unit (GPU) that NVIDIA designed for the Chinese market to comply with United States export controls on advanced artificial-intelligence accelerators.
The NVIDIA H200 is a data center Tensor Core GPU for AI and high-performance computing that was the first GPU to ship with HBM3e memory, packing 141 GB at 4.8 TB/s on its SXM and NVL boards.
The NVIDIA H800 is a data center graphics processing unit (GPU) that Nvidia designed specifically for the Chinese market as a regulatory-compliant variant of its flagship H100 accelerator.
NVIDIA HGX is a family of accelerated-server platform designs built around tightly connected data-center GPUs. It is not one immutable board specification.
NVIDIA Halos is a full-stack, comprehensive safety system developed by NVIDIA that unifies AI compute and safety across silicon, systems, software, and tools and services.
NVIDIA Holoscan is a domain-agnostic, multimodal AI sensor processing platform and software development kit (SDK) built by NVIDIA for real-time
NVIDIA Hopper is the codename for NVIDIA's ninth-generation datacenter GPU microarchitecture, announced March 22, 2022 by CEO Jensen Huang at the GTC keynote.
NVIDIA Isaac Lab is an open-source, GPU-accelerated framework for robot learning that trains robot control policies at scale by running thousands of physics simulations in parallel on a single graphics card
NVIDIA Isaac Lab-Arena is an open-source framework for composing simulated robot tasks and evaluating learned policies at scale.
NVIDIA Isaac Sim is an open-source robotics simulation application built on NVIDIA Omniverse for developing, simulating, and testing AI-driven robots in physically based virtual environments.
NVIDIA Isaac for Healthcare is NVIDIA's open-source developer framework for building, simulating, training, and deploying medical robots.
NVIDIA Ising is a family of open AI models from NVIDIA for operating quantum computers, covering two of the field's main engineering bottlenecks: quantum processor calibration and quantum error-correction…
NVIDIA Jetson is NVIDIA's family of compact, power-efficient system-on-modules (SoMs) that bring GPU-accelerated computing to machines that cannot lean on a data-center connection: robots, drones, cameras, and…
NVIDIA Jetson T4000 is a commercial system-on-module developed by Nvidia for edge AI and robotics. It is a member of the NVIDIA Jetson Thor family and became available on January 5, 2026.
The NVIDIA L4 is a compact, power-efficient data-center GPU built on the Ada Lovelace architecture and optimized for artificial-intelligence inference and video processing.
The NVIDIA L40S is a dual-slot, passively cooled data-center GPU based on the Ada Lovelace architecture.
NVIDIA MGX is a modular reference architecture that NVIDIA publishes so that server makers and contract manufacturers can build accelerated systems around NVIDIA GPUs, CPUs, DPUs and networking without…
NVIDIA NIM (NVIDIA Inference Microservices) is a set of containerized, prebuilt-and-optimized model-serving microservices from NVIDIA that package an AI model, an optimized inference engine, and an…
NVIDIA NeMo is an open, end-to-end framework from NVIDIA for building, customizing, and deploying generative AI models, described by NVIDIA as "a scalable generative AI framework built for researchers and…
The NVIDIA NeMo Agent Toolkit is an open-source, framework-agnostic library for connecting, profiling, evaluating, and optimizing teams of AI agents.
NVIDIA NeMo Switchyard is an open-source proxy and Rust library for routing requests among configured large language models.
NVIDIA NemoClaw is a collection of open blueprints (reference architectures) for building custom autonomous AI agents that packages three moving parts, a model, an agent harness, and a secure runtime
The NVIDIA Nemotron Model Reasoning Challenge was a Kaggle competition run by NVIDIA from March to June 2026 in which participants tried to improve the reasoning accuracy of a fixed open model, Nemotron 3 Nano…
Newton is an open-source, GPU-accelerated physics engine built for robotics simulation and robot learning, co-developed by NVIDIA, Google DeepMind, and Disney Research and stewarded by the Linux Foundation.
NVIDIA OSMO is an open-source workflow orchestration platform developed by Nvidia for physical AI and robotics development.
NVIDIA Omniverse is a scalable, multi-GPU real-time 3D graphics collaboration and simulation platform developed by NVIDIA for building and operating industrial digital twins and physical AI applications.
NVIDIA OpenShell is an open source runtime that executes autonomous AI agents inside policy-governed sandboxes, published by NVIDIA under the Apache License 2.0.
Parakeet is a family of open automatic speech recognition (ASR) models developed by NVIDIA as part of the NeMo conversational AI toolkit.
NVIDIA Picasso is a cloud-based generative AI foundry from NVIDIA for building, training, and deploying visual generative models that produce images, video, and 3D content from text prompts.
NVIDIA Quantum-X Photonics is a family of co-packaged-optics network switches that NVIDIA announced at its GTC conference on March 18, 2025.
The NVIDIA RTX PRO 6000 Blackwell is a professional workstation and server graphics processing unit (GPU) built on NVIDIA's Blackwell architecture, pairing the large GB202 die with 96 GB of GDDR7 memory and…
NVIDIA RTX Spark is a consumer "superchip" for Windows personal computers, announced by NVIDIA on May 31, 2026, at the company's GTC Taipei keynote held during COMPUTEX 2026.
NVIDIA Riva is a GPU-accelerated software development kit and family of containerized inference services for speech and translation AI, built by NVIDIA.
NVIDIA Rubin CPX is a class of GPU announced by NVIDIA on September 9, 2025, purpose-built to accelerate the compute-heavy "context" phase of large-model inference.
NVIDIA Rubin Ultra is a planned data-center GPU platform from Nvidia, positioned on the company's roadmap as the mid-cycle "Ultra" refresh of the Rubin generation.
NVIDIA Spectrum-6 is an Ethernet switch ASIC and system architecture for large AI infrastructure.
NVIDIA Spectrum-X is an Ethernet networking platform built by NVIDIA specifically for large-scale AI data center workloads
NVIDIA Spectrum-X Photonics is a co-packaged optics (CPO) Ethernet switch platform from Nvidia, announced at the company's GPU Technology Conference (GTC) on March 18, 2025.
NVIDIA TensorRT-LLM is an open-source library developed by nvidia for high-performance inference of large language models on NVIDIA GPUs.
NVIDIA Triton Inference Server is open-source model deployment software that lets teams run trained models from any machine learning or deep learning framework on any processor (GPU, CPU, or other accelerator)…
NVIDIA Vera is a custom Arm-based data center central processing unit from NVIDIA, unveiled at the GTC Taipei keynote at COMPUTEX 2026 on June 1, 2026 and marketed by the company as "the CPU for agents," a…
NVIDIA Vera Rubin is NVIDIA's data center AI computing platform that succeeds the NVIDIA Blackwell architecture, pairing the custom Arm-based Vera CPU with the dual-die Rubin GPU into a single superchip built…
NVIDIA Warp is an open-source Python framework, first released by NVIDIA in March 2022, that takes ordinary Python functions and just-in-time compiles them into native kernels that run on the CPU or on a…
The NVIDIA acquisition of Hugging Face is a pending transaction in which NVIDIA agreed to buy Hugging Face, the New York company that operates the largest hosting platform for open machine-learning models and…
NVLM (short for NVIDIA Vision Language Model), released as NVLM 1.0, is a family of open multimodal large language models developed by Nvidia.
NVLink is NVIDIA's proprietary high-bandwidth, low-latency, cache-coherent point-to-point interconnect for directly connecting GPUs to other GPUs and, in some configurations, to host CPUs.
NVLink Fusion is a program and silicon technology from NVIDIA that opens its NVLink high speed interconnect to third party chips.
NVSwitch is a family of switch chips (ASICs) designed by Nvidia that fully connect multiple GPUs over NVLink into a single shared-memory fabric, letting every GPU talk to every other GPU at full link speed.
Nemotron is NVIDIA's brand for its family of open large language models and the datasets, training recipes, and evaluation tools built around them.
Nemotron 3 is a family of open-weights large language model systems released by NVIDIA beginning on December 15, 2025, built for agentic AI and consisting of three sparse mixture-of-experts variants named…