AI Hardware

Explore AI Hardware through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: NVIDIA

Articles that also belong to these categories. Counts cover all of AI Hardware.

Showing 1-60 of 61 articles

Ada Lovelace (microarchitecture)

Ada Lovelace is the graphics processing unit (GPU) microarchitecture that Nvidia announced on September 20, 2022, succeeding the consumer side of the Ampere architecture and serving as the foundation for the…

NVIDIA

Ampere (microarchitecture)

Ampere is a graphics processing unit (GPU) microarchitecture from Nvidia, announced on May 14, 2020, as the successor to the Volta and Turing architectures.

NVIDIA

ChipNeMo

ChipNeMo is a research project and a family of domain-adapted large language models developed by Nvidia to assist with industrial semiconductor and chip-design tasks.

Large Language ModelsNVIDIA

CuDNN

NVIDIA cuDNN (CUDA Deep Neural Network library) is a proprietary GPU-accelerated library of primitives for deep learning, first released by NVIDIA on September 7, 2014, that provides highly tuned…

AI Tools & ProductsDeveloper Tools

Jensen Huang

Jen-Hsun "Jensen" Huang (born 1963) is a Taiwan-born electrical engineer and business executive based in the United States.

NVIDIAPeople

Jetson Thor

NVIDIA Jetson Thor (also marketed as Jetson AGX Thor) is a Blackwell architecture edge AI computing module developed by NVIDIA that delivers up to 2,070 FP4 teraflops of AI compute in a 40 to 130 watt power…

NVIDIARobotics

Mellanox

Mellanox Technologies, Ltd. was an Israeli-American semiconductor and networking company, founded in 1999, that designed the high-performance interconnects, network adapters, switches, cables, and programmable…

AI CompaniesNVIDIA

NVFP4

NVFP4 (NVIDIA FP4) is a 4-bit floating-point number format introduced by Nvidia with the Blackwell GPU architecture.

Machine LearningNVIDIA

NVHBM

NVHBM is an announced custom high-bandwidth memory architecture from NVIDIA for custom AI accelerators that participate in the company's NVLink Fusion platform. NVIDIA introduced it on August 26, 2026.

AI InfrastructureNVIDIA

NVIDIA A100

The NVIDIA A100 Tensor Core GPU is a datacenter graphics processing unit that NVIDIA introduced on May 14, 2020 as the first product built on its Ampere architecture, and that became the dominant chip for…

Data CentersNVIDIA

NVIDIA A800

The NVIDIA A800 is a datacenter graphics processing unit that Nvidia created for the Chinese market in late 2022 as an export-compliant variant of the A100

Chinese AINVIDIA

NVIDIA B100

The NVIDIA B100 is a data center graphics processing unit (GPU) based on the Blackwell architecture, announced by NVIDIA CEO Jensen Huang at the GTC 2024 keynote on March 18, 2024.

NVIDIA

NVIDIA Blackwell B200

The NVIDIA B200 is a data center GPU accelerator built on the Blackwell microarchitecture, introduced by NVIDIA on March 18, 2024 at the company's GTC keynote in San Jose.

NVIDIA

NVIDIA Blackwell Ultra

NVIDIA Blackwell Ultra is a mid-cycle refresh of NVIDIA's Blackwell data-center GPU architecture, announced at NVIDIA's GTC conference on March 18, 2025, and built for what NVIDIA calls the "age of AI…

NVIDIA

NVIDIA BlueField

NVIDIA BlueField is a family of data processing units (DPUs) designed and sold by Nvidia that offload and accelerate networking, storage, and security tasks away from a server's main CPU.

NVIDIA

NVIDIA ConnectX

NVIDIA ConnectX is a family of high-speed network adapters and SmartNICs (smart network interface cards) that connect a server to the data center fabric and accelerate networking in hardware.

AI InfrastructureNVIDIA

NVIDIA DGX Spark

NVIDIA DGX Spark is a compact, deskside AI development system, marketed by NVIDIA as a "personal AI supercomputer," that puts a Grace Blackwell superchip on a desktop and delivers up to 1 petaFLOP (1,000 TOPS)…

NVIDIA

NVIDIA DGX Station

NVIDIA DGX Station is a line of deskside artificial intelligence workstations from NVIDIA, each marketed as a "personal AI supercomputer" that puts data-center-class compute next to a developer's desk rather…

AI InfrastructureNVIDIA

NVIDIA DGX Station for Windows

NVIDIA DGX Station for Windows is a deskside artificial intelligence supercomputer announced by NVIDIA on June 1, 2026, at NVIDIA GTC Taipei during COMPUTEX 2026.

NVIDIA

NVIDIA DSX

NVIDIA DSX is a platform from NVIDIA for designing, simulating, building and operating large-scale data centers that the company calls AI factories.

Data CentersNVIDIA

NVIDIA Feynman

NVIDIA Feynman is the data center GPU architecture that NVIDIA has placed on its public roadmap as the successor to Rubin and Rubin Ultra, with a target of around 2028.

NVIDIA

NVIDIA GB200 NVL72

The NVIDIA GB200 NVL72 is a rack-scale AI computing system that packages 72 Blackwell GPUs and 36 Grace CPUs into a single liquid-cooled rack

NVIDIA

NVIDIA GB300 NVL72

The NVIDIA GB300 NVL72 is a liquid-cooled, rack-scale AI computing system that integrates 72 NVIDIA Blackwell Ultra (B300) GPUs and 36 Arm-based NVIDIA Grace CPUs into a single NVLink fabric, delivering 1.1…

AI InfrastructureData Centers

NVIDIA Grace

NVIDIA Grace is an Arm-based data-center central processing unit (CPU) from Nvidia, built from 72 Arm Neoverse V2 cores with co-packaged LPDDR5X memory and a coherent NVLink-C2C interconnect that links it to…

NVIDIA

NVIDIA Groq LPX Rack

NVIDIA Groq 3 LPX is a rack-scale inference accelerator that NVIDIA introduced at GTC 2026, built around 256 Groq Language Processing Units and designed to sit beside Vera Rubin NVL72 racks as a dedicated…

AI InferenceNVIDIA

NVIDIA H100

NVIDIA H100 (also called the H100 Tensor Core GPU) is a data-center graphics processing unit built by NVIDIA on the Hopper microarchitecture, fabricated with over 80 billion transistors on a custom TSMC 4N (4…

AI InfrastructureNVIDIA

NVIDIA H20

The NVIDIA H20 is a data-center graphics processing unit (GPU) that NVIDIA designed for the Chinese market to comply with United States export controls on advanced artificial-intelligence accelerators.

Chinese AINVIDIA

NVIDIA H200

The NVIDIA H200 is a data center Tensor Core GPU for AI and high-performance computing that was the first GPU to ship with HBM3e memory, packing 141 GB at 4.8 TB/s on its SXM and NVL boards.

Data CentersNVIDIA

NVIDIA H800

The NVIDIA H800 is a data center graphics processing unit (GPU) that Nvidia designed specifically for the Chinese market as a regulatory-compliant variant of its flagship H100 accelerator.

Chinese AINVIDIA

NVIDIA Hopper

NVIDIA Hopper is the codename for NVIDIA's ninth-generation datacenter GPU microarchitecture, announced March 22, 2022 by CEO Jensen Huang at the GTC keynote.

Data CentersNVIDIA

NVIDIA Jetson

NVIDIA Jetson is NVIDIA's family of compact, power-efficient system-on-modules (SoMs) that bring GPU-accelerated computing to machines that cannot lean on a data-center connection: robots, drones, cameras, and…

NVIDIA

NVIDIA Jetson T4000

NVIDIA Jetson T4000 is a commercial system-on-module developed by Nvidia for edge AI and robotics. It is a member of the NVIDIA Jetson Thor family and became available on January 5, 2026.

NVIDIARobotics

NVIDIA L4

The NVIDIA L4 is a compact, power-efficient data-center GPU built on the Ada Lovelace architecture and optimized for artificial-intelligence inference and video processing.

NVIDIA

NVIDIA L40S

The NVIDIA L40S is a dual-slot, passively cooled data-center GPU based on the Ada Lovelace architecture.

NVIDIA

NVIDIA MGX

NVIDIA MGX is a modular reference architecture that NVIDIA publishes so that server makers and contract manufacturers can build accelerated systems around NVIDIA GPUs, CPUs, DPUs and networking without…

AI InfrastructureNVIDIA

NVIDIA Picasso

NVIDIA Picasso is a cloud-based generative AI foundry from NVIDIA for building, training, and deploying visual generative models that produce images, video, and 3D content from text prompts.

AI InferenceAI Infrastructure

NVIDIA RTX PRO 6000 Blackwell

The NVIDIA RTX PRO 6000 Blackwell is a professional workstation and server graphics processing unit (GPU) built on NVIDIA's Blackwell architecture, pairing the large GB202 die with 96 GB of GDDR7 memory and…

NVIDIA

NVIDIA RTX Spark

NVIDIA RTX Spark is a consumer "superchip" for Windows personal computers, announced by NVIDIA on May 31, 2026, at the company's GTC Taipei keynote held during COMPUTEX 2026.

NVIDIA

NVIDIA Rubin CPX

NVIDIA Rubin CPX is a class of GPU announced by NVIDIA on September 9, 2025, purpose-built to accelerate the compute-heavy "context" phase of large-model inference.

AI InferenceNVIDIA

NVIDIA Rubin Ultra

NVIDIA Rubin Ultra is a planned data-center GPU platform from Nvidia, positioned on the company's roadmap as the mid-cycle "Ultra" refresh of the Rubin generation.

NVIDIA

NVIDIA Vera (CPU)

NVIDIA Vera is a custom Arm-based data center central processing unit from NVIDIA, unveiled at the GTC Taipei keynote at COMPUTEX 2026 on June 1, 2026 and marketed by the company as "the CPU for agents," a…

NVIDIA

NVIDIA Vera Rubin

NVIDIA Vera Rubin is NVIDIA's data center AI computing platform that succeeds the NVIDIA Blackwell architecture, pairing the custom Arm-based Vera CPU with the dual-die Rubin GPU into a single superchip built…

NVIDIA

NVLink

NVLink is NVIDIA's proprietary high-bandwidth, low-latency, cache-coherent point-to-point interconnect for directly connecting GPUs to other GPUs and, in some configurations, to host CPUs.

Data CentersNVIDIA

NVSwitch

NVSwitch is a family of switch chips (ASICs) designed by Nvidia that fully connect multiple GPUs over NVLink into a single shared-memory fabric, letting every GPU talk to every other GPU at full link speed.

NVIDIA

SpaceX Starmind

SpaceX Starmind is SpaceX's planned constellation of solar-powered artificial intelligence compute satellites, intended to function as orbital data centers that run AI workloads in space and beam results back…

AI InfrastructureData Centers

Tensor Core

A Tensor Core is a specialized execution unit inside NVIDIA GPUs that computes a small matrix multiplication and accumulation, D = A x B + C

AI InfrastructureNVIDIA

Volta (microarchitecture)

Volta is a GPU microarchitecture developed by Nvidia and introduced in 2017. It is best known as the architecture of the Tesla V100, the data center accelerator that brought the first generation of Tensor…

NVIDIA