AI Accelerator Comparison (H100 vs B200 vs MI300 vs TPU)
As of July 2026, the best AI accelerator depends on the metric you care about, so the honest answer to "H100 vs B200 vs MI300 vs TPU" is not a single winner.
Explore Data Centers through related topics and the articles other pages reference most.
Ranked by links from other AI Wiki pages.
Articles that also belong to these categories. Counts cover all of Data Centers.
Showing 1-57 of 57 articles
As of July 2026, the best AI accelerator depends on the metric you care about, so the honest answer to "H100 vs B200 vs MI300 vs TPU" is not a single winner.
The AI data center moratorium wave of 2026 is a surge of local and state pauses, bans, and restrictive legislation aimed at hyperscale data center development across the United States.
The AMD Instinct MI300X is a data center GPU accelerator that Advanced Micro Devices released on December 6, 2023, built on the CDNA 3 architecture and paired with 192 GB of HBM3 memory at 5.3 TB/s of bandwidth
The AMD Instinct MI325X is a data center GPU accelerator from AMD for AI training and inference, built on the CDNA 3 architecture with 256 GB of HBM3E memory and 6 TB/s of memory bandwidth
The AMD Instinct MI355X is a data center GPU accelerator built on AMD's CDNA 4 architecture, announced at the AMD Advancing AI 2025 event on June 12, 2025, and reaching general availability in October 2025.
AMD Pensando is the data processing unit (DPU) and AI networking product line of AMD, built around Pensando Systems, a startup AMD acquired in 2022 in a transaction valued at approximately $1.9 billion .
AWS Trainium 2 (also written as Trainium2 and abbreviated Trn2) is the second generation of Amazon Web Services' custom machine learning training accelerator
The Anthropic-Amazon Trainium expansion is an enlarged compute and investment agreement between Anthropic and Amazon, announced on 20 April 2026, under which Amazon committed to invest up to roughly $25…
CME Compute Futures are a planned family of financially settled futures contracts tied to Silicon Data benchmarks for hourly rentals of Nvidia H100 and B200 graphics processors.
Catalina is a high-power, liquid-cooled rack system designed by Meta for training and serving large AI models.
The Cerebras WSE-3 (Wafer-Scale Engine 3) is the third-generation wafer-scale AI chip developed by Cerebras Systems, announced on March 13, 2024, and is the largest semiconductor ever built.
Cloud AI GPU pricing is difficult to compare from a headline hourly rate alone. Providers sell different GPU variants, node sizes, CPU and memory bundles, regions, network configurations, and purchase models.
ClusterMAX is a rating and ranking system for GPU cloud providers published by the research firm SemiAnalysis.
Co-packaged optics (CPO) is a hardware architecture that places active optical engines on the same first-level package substrate as a host application-specific integrated circuit (ASIC).
Crusoe is an American AI infrastructure company headquartered in Denver, Colorado, that builds and operates data centers for artificial intelligence workloads.
Crusoe Energy Systems, Inc. (commonly branded as Crusoe) is a privately held American energy and computing infrastructure company headquartered in Denver, Colorado.
Ethernet is the family of wired computer networking technologies standardized by the IEEE 802.3 working group.
A GPU cluster is a group of servers containing graphics processing units that are connected by high-bandwidth, low-latency networks and operated as one computational pool.
Virgo Network is a megascale data-center fabric introduced by Google at Google Cloud Next 2026 in April 2026.
Grand Teton is an open GPU hardware platform designed by Meta for training and running large AI models.
Hyperion is an artificial intelligence data center campus built by Meta in Richland Parish, in northeast Louisiana, United States.
A hyperscaler is a company that builds and operates computing infrastructure at a scale far beyond a conventional enterprise IT estate: globally distributed fleets of data centers holding millions of servers…
Land, power, and shell (LPS) is an emerging commercial label for the physical site, deliverable electricity, and building structure needed before computing equipment can be installed in a large data center.
Meta Compute is a top-level organization that Meta created in January 2026 to plan, build, and run the gigawatt-scale data center capacity behind its push toward artificial general intelligence and what the…
The Meta-Amazon Graviton deal is a multibillion-dollar, multiyear agreement announced on April 24, 2026
Fairwater is Microsoft's name for a class of large AI datacenter built to train and serve frontier artificial intelligence models.
Microsoft's microfluidic cooling is an experimental chip-cooling technique, disclosed by Microsoft in September 2025, that etches tiny channels directly into the back of a silicon die and pushes liquid coolant…
The NVIDIA A100 Tensor Core GPU is a datacenter graphics processing unit that NVIDIA introduced on May 14, 2020 as the first product built on its Ampere architecture, and that became the dominant chip for…
The NVIDIA AI compute infrastructure financing platforms are a set of proposed, independently run financing vehicles that NVIDIA announced on August 10, 2026, together with Apollo, BlackRock, Blackstone…
The NVIDIA B200 is a data center GPU based on the NVIDIA Blackwell microarchitecture, announced by Jensen Huang at GTC 2024 on March 18, 2024.
The NVIDIA DGX B300 is an 8-GPU AI supercomputer node built around NVIDIA's Blackwell Ultra architecture.
The NVIDIA DGX SuperPOD is a reference-architecture artificial intelligence supercomputer designed and sold by Nvidia.
NVIDIA DSX is a platform from NVIDIA for designing, simulating, building and operating large-scale data centers that the company calls AI factories.
NVIDIA Exemplar Cloud is a validation program run by NVIDIA that certifies cloud providers whose GPU clusters reproduce at least 95% of the training throughput NVIDIA measures on its own reference architecture…
The NVIDIA GB300 NVL72 is a liquid-cooled, rack-scale AI computing system that integrates 72 NVIDIA Blackwell Ultra (B300) GPUs and 36 Arm-based NVIDIA Grace CPUs into a single NVLink fabric, delivering 1.1…
The NVIDIA H200 is a data center Tensor Core GPU for AI and high-performance computing that was the first GPU to ship with HBM3e memory, packing 141 GB at 4.8 TB/s on its SXM and NVL boards.
NVIDIA HGX is a family of accelerated-server platform designs built around tightly connected data-center GPUs. It is not one immutable board specification.
NVIDIA Hopper is the codename for NVIDIA's ninth-generation datacenter GPU microarchitecture, announced March 22, 2022 by CEO Jensen Huang at the GTC keynote.
NVIDIA Spectrum-6 is an Ethernet switch ASIC and system architecture for large AI infrastructure.
NVIDIA Spectrum-X is an Ethernet networking platform built by NVIDIA specifically for large-scale AI data center workloads
NVLink is NVIDIA's proprietary high-bandwidth, low-latency, cache-coherent point-to-point interconnect for directly connecting GPUs to other GPUs and, in some configurations, to host CPUs.
A neocloud is a cloud computing company whose business is renting out GPU capacity for artificial intelligence workloads, rather than selling the broad catalogue of storage, database, networking and…
PCI Express (PCIe) is a high-speed serial interconnect standard used to attach processors to graphics cards, network adapters, storage devices, and accelerators inside nearly every modern computer and server.
The PORTS Technology Campus, also called the PORTS-Pike Technology Campus, is a planned AI infrastructure and power-generation complex in Pike County, Ohio.
Project Rainier is a distributed artificial intelligence supercomputer built by Amazon Web Services around its in-house AWS Trainium accelerators, created principally to train and serve the Claude models of…
Prometheus is a roughly 1-gigawatt AI supercluster built by Meta at its data center campus in New Albany, Ohio.
Qualcomm AI200 is a rack-scale data-center accelerator for artificial intelligence inference, announced by Qualcomm on 27 October 2025 and slated for commercial availability in 2026 .
Qualcomm AI250 is a planned data-center artificial intelligence inference accelerator and rack-scale system announced by Qualcomm in late October 2025.
The Research SuperCluster (RSC) is an AI supercomputer built by Meta AI, the artificial intelligence research division of Meta Platforms (the company formerly known as Facebook).
Space-based data centers are data centers launched into Earth orbit (or, in some proposals, placed on the Moon), where satellites carrying AI accelerators draw power from solar arrays and reject heat by…
SpaceX Starmind is SpaceX's planned constellation of solar-powered artificial intelligence compute satellites, intended to function as orbital data centers that run AI workloads in space and beam results back…
Starcloud, Inc. is an American space infrastructure startup that builds space-based data centers: satellites carrying data-center-class GPUs that are powered by solar arrays and cooled by radiating waste heat…
Stargate Michigan is an artificial intelligence data center campus under construction in Saline Township, Michigan, about 10 miles southwest of Ann Arbor in Washtenaw County.
Stargate UAE is a large artificial intelligence data center cluster being built in Abu Dhabi, United Arab Emirates, and is the first site of the Stargate Project located outside the United States.
TPU Ironwood (officially TPU v7 or TPU7x) is Google's seventh-generation Tensor Processing Unit and the first TPU designed specifically for inference, unveiled at Google Cloud Next 2025 in Las Vegas on April…
Voltage Park is a United States cloud computing company that operates a fleet of NVIDIA H100 graphics processing units for artificial intelligence training and inference workloads.
Colossus is an artificial-intelligence supercomputer and GPU training cluster operated by xAI in Memphis, Tennessee, and is widely described as one of the largest single-site AI GPU clusters in the world.