NVIDIA
113 AI Wiki articles on NVIDIA. The most referenced are CUDA, NVIDIA H100 and NVIDIA Blackwell.
113 articlesRSS
Showing 1-60 of 113 articles
Ada Lovelace (microarchitecture)
Ada Lovelace is the graphics processing unit (GPU) microarchitecture that Nvidia announced on September 20, 2022, succeeding the consumer side of the Ampere...
AI Hardware
Ampere (microarchitecture)
Ampere is a graphics processing unit (GPU) microarchitecture from Nvidia, announced on May 14, 2020, as the successor to the Volta and Turing architectures. It...
AI Hardware
CUDA
CUDA (Compute Unified Device Architecture) is NVIDIA's platform and programming model for general-purpose computation on its graphics processing units. NVIDIA...
AI InfrastructureDeveloper Tools
CUTLASS
CUTLASS is an open-source library of reusable building blocks for writing high-performance matrix kernels on NVIDIA GPUs. The name expands to CUDA Templates...
AI InfrastructureDeveloper Tools
ChipNeMo
ChipNeMo is a research project and a family of domain-adapted large language models developed by Nvidia to assist with industrial semiconductor and chip-design...
AI HardwareLarge Language Models
CuDNN
NVIDIA cuDNN (CUDA Deep Neural Network library) is a proprietary GPU-accelerated library of primitives for deep learning, first released by NVIDIA on September...
AI HardwareAI Tools & Products
Deepu Talla
Deepu Talla is the vice president of Robotics and Edge AI at NVIDIA. He leads an organization that brings artificial intelligence to autonomous machines and...
PeopleRobotics
EDM (Elucidating Diffusion Models)
EDM is the common shorthand for the paper "Elucidating the Design Space of Diffusion-Based Generative Models" by Tero Karras, Miika Aittala, Timo Aila, and...
Diffusion ModelsGenerative AI
GPU Technology Conference
The GPU Technology Conference (GTC) is the annual flagship technology conference hosted by Nvidia, and it is the principal stage on which Nvidia reveals new...
AI Events
Isaac GR00T
Isaac GR00T (short for Generalist Robot 00 Technology) is an open foundation model family from NVIDIA that gives humanoid robots a single, pretrained "AI...
Embodied AIHumanoid Robots
Jensen Huang
Jen-Hsun "Jensen" Huang (born 1963) is a Taiwan-born electrical engineer and business executive based in the United States. He co-founded NVIDIA in 1993 and...
AI HardwarePeople
Jet-Nemotron
Jet-Nemotron is a family of small hybrid-architecture language models released by NVIDIA Research in August 2025.[1] The family ships in two sizes,...
AI ModelsAI Research
Jetson Thor
NVIDIA Jetson Thor (also marketed as Jetson AGX Thor) is a Blackwell architecture edge AI computing module developed by NVIDIA that delivers up to 2,070 FP4...
AI HardwareRobotics
KAI Scheduler
KAI Scheduler is an open-source Kubernetes scheduler that optimizes the allocation of GPU resources for artificial intelligence and machine learning workloads....
AI InfrastructureMLOps
Lepton AI
Subsidiary of NVIDIA (since 2025); formerly private company Industry 2023 Founders Cupertino / San Francisco Bay Area, California, United States Key...
AI CompaniesAI Infrastructure
Llama Nemotron
Llama Nemotron is a family of open reasoning large language models built by Nvidia by post-training Meta's Llama models for math, coding, and agentic tasks....
Large Language ModelsReasoning Models
Llama-3.1-Nemotron-70B-Instruct
Llama-3.1-Nemotron-70B-Instruct is a large language model released by NVIDIA in October 2024. It is a customized, alignment-tuned version of Meta's Llama 3.1...
AI ModelsLarge Language Models
Megatron-LM
Megatron-LM is NVIDIA's open-source framework for training very large transformer language models across GPU clusters, and the name of the tensor-parallelism...
Open Source AITraining & Optimization
Mellanox
Mellanox Technologies, Ltd. was an Israeli-American semiconductor and networking company, founded in 1999, that designed the high-performance interconnects,...
AI CompaniesAI Hardware
MimicGen
MimicGen is a data generation system developed by researchers at NVIDIA's Seattle Robotics Lab and Learning and Perception Research group that automatically...
Data & DatasetsEmbodied AI
Minitron
Minitron is a family of compact language models from NVIDIA, together with the model-compression method used to build them: take one large, already pretrained...
Small Language ModelsTraining & Optimization
Mistral NeMo
Mistral NeMo is a 12 billion parameter large language model released by Mistral AI in collaboration with NVIDIA on July 18, 2024 [1][2]. It shipped under the...
Large Language ModelsOpen Source AI
NCCL (NVIDIA Collective Communications Library)
NCCL, the NVIDIA Collective Communications Library, is a library of topology-aware communication primitives for NVIDIA GPU systems. It supplies collective...
AI HardwareAI Infrastructure
NOOA (NVIDIA Object-Oriented Agents)
NOOA (NVIDIA Object-Oriented Agents) is an open-source, model-agnostic Python framework for building AI agents, released by NVIDIA in July 2026 [1][2]. Its...
AI AgentsDeveloper Tools
NVFP4
NVFP4 (NVIDIA FP4) is a 4-bit floating-point number format introduced by Nvidia with the Blackwell GPU architecture. It stores each value in just 4 bits using...
AI HardwareMachine Learning
NVIDIA A100
The NVIDIA A100 Tensor Core GPU is a datacenter graphics processing unit that NVIDIA introduced on May 14, 2020 as the first product built on its Ampere...
AI HardwareData Centers
NVIDIA A800
The NVIDIA A800 is a datacenter graphics processing unit that Nvidia created for the Chinese market in late 2022 as an export-compliant variant of the A100,...
AI HardwareChinese AI
NVIDIA AI Enterprise
NVIDIA AI Enterprise is an end-to-end, cloud-native software suite sold by Nvidia as a paid subscription for developing and deploying production artificial...
AI InfrastructureEnterprise AI
NVIDIA Alpamayo 2 Super
NVIDIA Alpamayo 2 Super is an open, 34-billion-parameter reasoning-based vision-language-action model (VLA) for safe, Level 4 robotaxi and autonomous-vehicle...
AI ModelsAutonomous Vehicles
NVIDIA B100
The NVIDIA B100 is a data center graphics processing unit (GPU) based on the Blackwell architecture, announced by NVIDIA CEO Jensen Huang at the GTC 2024...
AI Hardware
NVIDIA B200
The NVIDIA B200 is a data center GPU based on the NVIDIA Blackwell microarchitecture, announced by Jensen Huang at GTC 2024 on March 18, 2024.[^1] It is the...
AI HardwareAI Infrastructure
NVIDIA Blackwell
!Nvidia blackwell1.jpg !Nvidia blackwell2.jpg NVIDIA Blackwell is a family of graphics processing unit architectures and computing platforms developed by...
AI HardwareDeep Learning
NVIDIA Blackwell B200
The NVIDIA B200 is a data center GPU accelerator built on the Blackwell microarchitecture, introduced by NVIDIA on March 18, 2024 at the company's GTC keynote...
AI Hardware
NVIDIA Blackwell Ultra
NVIDIA Blackwell Ultra is a mid-cycle refresh of NVIDIA's Blackwell data-center GPU architecture, announced at NVIDIA's GTC conference on March 18, 2025, and...
AI Hardware
NVIDIA BlueField
NVIDIA BlueField is a family of data processing units (DPUs) designed and sold by Nvidia that offload and accelerate networking, storage, and security tasks...
AI Hardware
NVIDIA Canary
Canary is a family of open speech models developed by Nvidia as part of its NeMo conversational AI toolkit. Canary models perform both automatic speech...
AI ModelsSpeech & Audio AI
NVIDIA ConnectX
NVIDIA ConnectX is a family of high-speed network adapters and SmartNICs (smart network interface cards) that connect a server to the data center fabric and...
AI HardwareAI Infrastructure
NVIDIA Cosmos
NVIDIA Cosmos is a world foundation model platform developed by NVIDIA for physical AI applications, including autonomous vehicles and robotics. Announced by...
AI ModelsEmbodied AI
NVIDIA Cosmos 3
NVIDIA Cosmos 3 is an open family of "world foundation models" for physical AI that NVIDIA launched on June 1, 2026 at GTC Taipei, held alongside COMPUTEX...
Generative AIRobotics
NVIDIA Cosmos Reason
The original NVIDIA Cosmos Reason release is an open, customizable, 7-billion-parameter reasoning vision-language model (VLM) for physical AI and robotics...
Embodied AIMultimodal AI
NVIDIA DGX B300
The NVIDIA DGX B300 is an 8-GPU AI supercomputer node built around NVIDIA's Blackwell Ultra architecture. Announced at GTC 2025 in March[4] and shipping from...
AI HardwareAI Infrastructure
NVIDIA DGX Cloud
NVIDIA DGX Cloud is a managed AI-supercomputing-as-a-service offering from Nvidia that rents enterprises access to multi-node clusters of NVIDIA DGX...
AI Infrastructure
NVIDIA DGX Spark
NVIDIA DGX Spark is a compact, deskside AI development system, marketed by NVIDIA as a "personal AI supercomputer," that puts a Grace Blackwell superchip on a...
AI Hardware
NVIDIA DGX Station
Value --- Deskside AI workstation, marketed as a "personal AI supercomputer" Manufacturer May 2017, at GTC 2017 Generations GB300 Grace Blackwell...
AI HardwareAI Infrastructure
NVIDIA DGX Station for Windows
NVIDIA DGX Station for Windows is a deskside artificial intelligence supercomputer announced by NVIDIA on June 1, 2026, at NVIDIA GTC Taipei during COMPUTEX...
AI Hardware
NVIDIA DGX SuperPOD
The NVIDIA DGX SuperPOD is a reference-architecture artificial intelligence supercomputer designed and sold by Nvidia. It is a turnkey, factory-validated...
AI InfrastructureData Centers
NVIDIA DRIVE Hyperion
NVIDIA DRIVE Hyperion is a production-ready reference architecture for autonomous vehicles developed by NVIDIA. It packages a validated in-vehicle compute...
AI HardwareAutonomous Vehicles
NVIDIA DRIVE Thor
NVIDIA DRIVE Thor (marketed as DRIVE AGX Thor) is a centralized automotive and robotics system-on-a-chip (SoC) developed by Nvidia. It is designed to...
AI HardwareAutonomous Vehicles
NVIDIA DSX
NVIDIA DSX is a platform from NVIDIA for designing, simulating, building and operating large-scale data centers that the company calls AI factories. NVIDIA...
AI HardwareData Centers
NVIDIA Deep Learning Institute
NVIDIA Deep Learning Institute (DLI) is the training and education arm of NVIDIA, offering hands-on courses, instructor-led workshops, and professional...
AI InfrastructureDeveloper Tools
NVIDIA DeepStream
NVIDIA DeepStream is a GPU-accelerated software development kit for building real-time streaming analytics pipelines, principally for video. It is built on the...
Computer VisionDeveloper Tools
NVIDIA Dynamo
NVIDIA Dynamo is an open-source, low-latency distributed inference serving framework designed to deploy and scale generative AI and reasoning models across...
AI InferenceAI Infrastructure
NVIDIA Exemplar Cloud
NVIDIA Exemplar Cloud is a validation program run by NVIDIA that certifies cloud providers whose GPU clusters reproduce at least 95% of the training throughput...
AI InfrastructureData Centers
NVIDIA Feynman
NVIDIA Feynman is the data center GPU architecture that NVIDIA has placed on its public roadmap as the successor to Rubin and Rubin Ultra, with a target of...
AI Hardware
NVIDIA GB200 NVL72
The NVIDIA GB200 NVL72 is a rack-scale AI computing system that packages 72 Blackwell GPUs and 36 Grace CPUs into a single liquid-cooled rack, joined into one...
AI Hardware
NVIDIA GB300 NVL72
The NVIDIA GB300 NVL72 is a liquid-cooled, rack-scale AI computing system that integrates 72 NVIDIA Blackwell Ultra (B300) GPUs and 36 Arm-based NVIDIA Grace...
AI HardwareAI Infrastructure
NVIDIA GH200 Grace Hopper Superchip
The NVIDIA GH200 Grace Hopper Superchip is a single-module processor from NVIDIA that combines a 72-core Grace Arm CPU with a Hopper-generation H100-class GPU...
AI HardwareAI Infrastructure
NVIDIA Grace
NVIDIA Grace is an Arm-based data-center central processing unit (CPU) from Nvidia, built from 72 Arm Neoverse V2 cores with co-packaged LPDDR5X memory and a...
AI Hardware
NVIDIA Groq LPX Rack
NVIDIA Groq 3 LPX is a rack-scale inference accelerator that NVIDIA introduced at GTC 2026, built around 256 Groq Language Processing Units and designed to sit...
AI HardwareAI Inference
NVIDIA H100
NVIDIA H100 (also called the H100 Tensor Core GPU) is a data-center graphics processing unit built by NVIDIA on the Hopper microarchitecture, fabricated with...
AI HardwareAI Infrastructure