Microsoft Azure Maia 100
The Microsoft Azure Maia 100, often referred to as Maia 100, is a custom artificial intelligence accelerator designed by Microsoft for large-scale generative AI workloads running inside its Azure cloud.
Explore AI Hardware through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Hardware.
Showing 181-240 of 355 articles
The Microsoft Azure Maia 100, often referred to as Maia 100, is a custom artificial intelligence accelerator designed by Microsoft for large-scale generative AI workloads running inside its Azure cloud.
Microsoft HoloLens was a self-contained mixed reality headset built by Microsoft and sold in two generations.
Microsoft Maia 200 is a custom artificial intelligence accelerator designed by Microsoft for inference workloads in the Azure cloud.
Project Solara is a hardware and software platform that Microsoft unveiled at its Build 2026 developer conference on June 2, 2026.
Microsoft's microfluidic cooling is an experimental chip-cooling technique, disclosed by Microsoft in September 2025, that etches tiny channels directly into the back of a silicon die and pushes liquid coolant…
MobileNet is a family of efficient convolutional neural network (CNN) architectures developed by Google for mobile and edge AI applications.
Mobileye Global Inc. (Nasdaq: MBLY) is an Israeli technology company headquartered in Jerusalem that develops vision-based advanced driver-assistance systems (ADAS) and autonomous driving technology
Moore Threads (Chinese: 摩尔线程; Moore Threads Technology Co., Ltd.) is a Beijing-based Chinese designer of graphics processing units founded in October 2020 by Zhang Jianzhong, a former vice president and…
NAND flash memory is the non-volatile storage technology that holds the data in solid-state drives, memory cards, USB sticks, phones and the flash tiers of a modern data center.
NAURA Technology Group Co., Ltd. (北方华创科技集团股份有限公司, romanised as Beifang Huachuang and rendered in English sources as both "NAURA" and "Naura Technology Group") is China's largest maker of semiconductor…
NCCL, the NVIDIA Collective Communications Library, is a library of topology-aware communication primitives for NVIDIA GPU systems.
NVFP4 (NVIDIA FP4) is a 4-bit floating-point number format introduced by Nvidia with the Blackwell GPU architecture.
NVHBM is an announced custom high-bandwidth memory architecture from NVIDIA for custom AI accelerators that participate in the company's NVLink Fusion platform. NVIDIA introduced it on August 26, 2026.
The NVIDIA A100 Tensor Core GPU is a datacenter graphics processing unit that NVIDIA introduced on May 14, 2020 as the first product built on its Ampere architecture, and that became the dominant chip for…
The NVIDIA A800 is a datacenter graphics processing unit that Nvidia created for the Chinese market in late 2022 as an export-compliant variant of the A100
The NVIDIA B100 is a data center graphics processing unit (GPU) based on the Blackwell architecture, announced by NVIDIA CEO Jensen Huang at the GTC 2024 keynote on March 18, 2024.
The NVIDIA B200 is a data center GPU based on the NVIDIA Blackwell microarchitecture, announced by Jensen Huang at GTC 2024 on March 18, 2024.
NVIDIA Blackwell is a family of graphics processing unit architectures and computing platforms developed by NVIDIA.
The NVIDIA B200 is a data center GPU accelerator built on the Blackwell microarchitecture, introduced by NVIDIA on March 18, 2024 at the company's GTC keynote in San Jose.
NVIDIA Blackwell Ultra is a mid-cycle refresh of NVIDIA's Blackwell data-center GPU architecture, announced at NVIDIA's GTC conference on March 18, 2025, and built for what NVIDIA calls the "age of AI…
NVIDIA BlueField is a family of data processing units (DPUs) designed and sold by Nvidia that offload and accelerate networking, storage, and security tasks away from a server's main CPU.
NVIDIA ConnectX is a family of high-speed network adapters and SmartNICs (smart network interface cards) that connect a server to the data center fabric and accelerate networking in hardware.
NVIDIA DGX is nvidia's line of integrated artificial-intelligence supercomputers, purpose-built systems that package the company's highest-end data-center GPUs, CPUs, high-speed nvlink and nvswitch…
The NVIDIA DGX B300 is an 8-GPU AI supercomputer node built around NVIDIA's Blackwell Ultra architecture.
NVIDIA DGX Spark is a compact, deskside AI development system, marketed by NVIDIA as a "personal AI supercomputer," that puts a Grace Blackwell superchip on a desktop and delivers up to 1 petaFLOP (1,000 TOPS)…
NVIDIA DGX Station is a line of deskside artificial intelligence workstations from NVIDIA, each marketed as a "personal AI supercomputer" that puts data-center-class compute next to a developer's desk rather…
NVIDIA DGX Station for Windows is a deskside artificial intelligence supercomputer announced by NVIDIA on June 1, 2026, at NVIDIA GTC Taipei during COMPUTEX 2026.
NVIDIA DRIVE Hyperion is a production-ready reference architecture for autonomous vehicles developed by NVIDIA.
NVIDIA DRIVE Thor (marketed as DRIVE AGX Thor) is a centralized automotive and robotics system-on-a-chip (SoC) developed by Nvidia.
NVIDIA DSX is a platform from NVIDIA for designing, simulating, building and operating large-scale data centers that the company calls AI factories.
NVIDIA Digits (originally announced as Project DIGITS, later renamed DGX Spark) is a personal AI supercomputer developed by NVIDIA and co-designed with MediaTek.
NVIDIA Feynman is the data center GPU architecture that NVIDIA has placed on its public roadmap as the successor to Rubin and Rubin Ultra, with a target of around 2028.
The NVIDIA GB200 NVL72 is a rack-scale AI computing system that packages 72 Blackwell GPUs and 36 Grace CPUs into a single liquid-cooled rack
The NVIDIA GB300 NVL72 is a liquid-cooled, rack-scale AI computing system that integrates 72 NVIDIA Blackwell Ultra (B300) GPUs and 36 Arm-based NVIDIA Grace CPUs into a single NVLink fabric, delivering 1.1…
The NVIDIA GH200 Grace Hopper Superchip is a single-module processor from NVIDIA that combines a 72-core Grace Arm CPU with a Hopper-generation H100-class GPU on one package
NVIDIA Grace is an Arm-based data-center central processing unit (CPU) from Nvidia, built from 72 Arm Neoverse V2 cores with co-packaged LPDDR5X memory and a coherent NVLink-C2C interconnect that links it to…
NVIDIA Groq 3 LPX is a rack-scale inference accelerator that NVIDIA introduced at GTC 2026, built around 256 Groq Language Processing Units and designed to sit beside Vera Rubin NVL72 racks as a dedicated…
NVIDIA H100 (also called the H100 Tensor Core GPU) is a data-center graphics processing unit built by NVIDIA on the Hopper microarchitecture, fabricated with over 80 billion transistors on a custom TSMC 4N (4…
The NVIDIA H20 is a data-center graphics processing unit (GPU) that NVIDIA designed for the Chinese market to comply with United States export controls on advanced artificial-intelligence accelerators.
The NVIDIA H200 is a data center Tensor Core GPU for AI and high-performance computing that was the first GPU to ship with HBM3e memory, packing 141 GB at 4.8 TB/s on its SXM and NVL boards.
The NVIDIA H800 is a data center graphics processing unit (GPU) that Nvidia designed specifically for the Chinese market as a regulatory-compliant variant of its flagship H100 accelerator.
NVIDIA HGX is a family of accelerated-server platform designs built around tightly connected data-center GPUs. It is not one immutable board specification.
NVIDIA Hopper is the codename for NVIDIA's ninth-generation datacenter GPU microarchitecture, announced March 22, 2022 by CEO Jensen Huang at the GTC keynote.
NVIDIA Jetson is NVIDIA's family of compact, power-efficient system-on-modules (SoMs) that bring GPU-accelerated computing to machines that cannot lean on a data-center connection: robots, drones, cameras, and…
NVIDIA Jetson T4000 is a commercial system-on-module developed by Nvidia for edge AI and robotics. It is a member of the NVIDIA Jetson Thor family and became available on January 5, 2026.
The NVIDIA L4 is a compact, power-efficient data-center GPU built on the Ada Lovelace architecture and optimized for artificial-intelligence inference and video processing.
The NVIDIA L40S is a dual-slot, passively cooled data-center GPU based on the Ada Lovelace architecture.
NVIDIA MGX is a modular reference architecture that NVIDIA publishes so that server makers and contract manufacturers can build accelerated systems around NVIDIA GPUs, CPUs, DPUs and networking without…
NVIDIA Picasso is a cloud-based generative AI foundry from NVIDIA for building, training, and deploying visual generative models that produce images, video, and 3D content from text prompts.
NVIDIA Quantum-X Photonics is a family of co-packaged-optics network switches that NVIDIA announced at its GTC conference on March 18, 2025.
The NVIDIA RTX PRO 6000 Blackwell is a professional workstation and server graphics processing unit (GPU) built on NVIDIA's Blackwell architecture, pairing the large GB202 die with 96 GB of GDDR7 memory and…
NVIDIA RTX Spark is a consumer "superchip" for Windows personal computers, announced by NVIDIA on May 31, 2026, at the company's GTC Taipei keynote held during COMPUTEX 2026.
NVIDIA Rubin CPX is a class of GPU announced by NVIDIA on September 9, 2025, purpose-built to accelerate the compute-heavy "context" phase of large-model inference.
NVIDIA Rubin Ultra is a planned data-center GPU platform from Nvidia, positioned on the company's roadmap as the mid-cycle "Ultra" refresh of the Rubin generation.
NVIDIA Spectrum-6 is an Ethernet switch ASIC and system architecture for large AI infrastructure.
NVIDIA Spectrum-X is an Ethernet networking platform built by NVIDIA specifically for large-scale AI data center workloads
NVIDIA Spectrum-X Photonics is a co-packaged optics (CPO) Ethernet switch platform from Nvidia, announced at the company's GPU Technology Conference (GTC) on March 18, 2025.
NVIDIA Spectrum-XGS, marketed in full as NVIDIA Spectrum-XGS Ethernet, is a networking technology from NVIDIA designed to link several geographically separated data centers into a single
NVIDIA Vera is a custom Arm-based data center central processing unit from NVIDIA, unveiled at the GTC Taipei keynote at COMPUTEX 2026 on June 1, 2026 and marketed by the company as "the CPU for agents," a…
NVIDIA Vera Rubin is NVIDIA's data center AI computing platform that succeeds the NVIDIA Blackwell architecture, pairing the custom Arm-based Vera CPU with the dual-die Rubin GPU into a single superchip built…