NVIDIA RTX PRO 6000 Blackwell

10 min read
Updated
Suggest editHistoryTalk
RawGraph

Last edited

Fact-checked

In review queue

Sources

10 citations

Revision

v2 · 1,965 words

Fact-checks are independent of edits: a reviewer re-verifies the article against its sources and stamps the date. How we verify

The NVIDIA RTX PRO 6000 Blackwell is a professional workstation and server graphics processing unit (GPU) built on NVIDIA's Blackwell architecture, pairing the large GB202 die with 96 GB of GDDR7 memory and 24,064 CUDA cores. Announced at the GPU Technology Conference (GTC) on March 18, 2025, it is the flagship of the RTX PRO Blackwell family of professional graphics products and the successor to the RTX 6000 Ada Generation. NVIDIA describes the Workstation Edition as "the most powerful desktop GPU ever created" and positions the lineup for agentic AI, generative AI, large-model inference, rendering, scientific simulation, and professional design and visualization.[1][5]

The RTX PRO 6000 Blackwell delivers a generational leap in memory capacity and AI throughput over its predecessor, doubling memory from 48 GB to 96 GB and adding support for low-precision FP4 inference. It is offered in three editions targeting desktop workstations, thermally constrained workstations, and rack-mounted servers. A modified, lower-bandwidth derivative reported as the RTX PRO 6000D was subsequently created for the Chinese market to comply with United States export controls.

What is the NVIDIA RTX PRO 6000 Blackwell?

The RTX PRO 6000 Blackwell is NVIDIA's top professional graphics card of the Blackwell generation, sold for desktop workstations, dense multi-GPU workstations, and enterprise servers. Unlike NVIDIA's pure data-center accelerators, it is a unified AI-and-graphics part: it carries fourth-generation RT Cores and DisplayPort outputs for ray-traced rendering and visualization while also providing fifth-generation Tensor Cores and 96 GB of memory for on-device and on-premises AI inference. The naming dropped the long-running "Quadro" branding and the prior generation's "RTX A" and "RTX Ada" schemes in favor of the "RTX PRO" family name, with the "6000" denoting the flagship tier.[1][5]

When was the RTX PRO 6000 Blackwell announced?

NVIDIA unveiled the RTX PRO Blackwell lineup on March 18, 2025, during GTC 2025 in San Jose, California. The announcement spanned desktop, laptop, and server form factors, with the RTX PRO 6000 Blackwell sitting at the top of the desktop and server stacks. Lower desktop tiers announced alongside it included the RTX PRO 5000, RTX PRO 4500, RTX PRO 4000, and RTX PRO 2000 Blackwell, while a separate laptop series ranged from the RTX PRO 5000 down to the RTX PRO 500 Blackwell.[1][2]

NVIDIA framed the launch around enabling professionals to "build and collaborate with agentic AI," emphasizing on-device and on-premises AI workloads, neural rendering, ray tracing, physical AI, simulation, and content creation. Bob Pette, NVIDIA's vice president of enterprise platforms, said: "Bringing NVIDIA Blackwell to workstations and servers will take productivity, performance and speed to new heights, accelerating AI inference serving, data science, visualization and content creation."[2]

The company named a broad ecosystem of system and cloud partners for the new GPUs, including Dell Technologies, Hewlett Packard Enterprise, Lenovo, Supermicro, Cisco, ASUS, GIGABYTE, BOXX, HP Inc., and PNY, along with cloud providers Amazon Web Services, Google Cloud, Microsoft Azure, CoreWeave, and Lambda.[2]

What editions are available?

The RTX PRO 6000 Blackwell is sold in three distinct editions that share the same GPU silicon and 96 GB of memory but differ in cooling, form factor, and power envelope.

EditionTargetCoolingForm factorMax power
RTX PRO 6000 Blackwell Workstation EditionDesktop workstationsActive, double flow-through5.4 in H x 12.0 in L, dual slot600 W
RTX PRO 6000 Blackwell Max-Q Workstation EditionMulti-GPU / thermally constrained workstationsActive4.4 in H x 10.5 in L, dual slot300 W
RTX PRO 6000 Blackwell Server EditionEnterprise servers and data centersPassive (requires chassis airflow)4.4 in H x 10.5 in L, dual slot400 W to 600 W (configurable)

The Workstation Edition is the full-power desktop card. The Max-Q Workstation Edition trades peak performance for a 300 W envelope, allowing denser multi-GPU configurations in a single workstation. The Server Edition is a passively cooled, data-center card that relies on host-chassis airflow, supports up to eight GPUs per server, and is offered in both air-cooled and liquid-cooled variants.[1][3][4]

What are its specifications?

All three editions are built on the GB202 die and expose the same core configuration of 24,064 CUDA cores, 752 fifth-generation Tensor Cores, and 188 fourth-generation RT Cores, paired with 96 GB of GDDR7 memory with error-correcting code (ECC) on a 512-bit interface. The fifth-generation Tensor Cores add support for the FP4 data format, and the Server Edition supports Multi-Instance GPU (MIG) partitioning into up to four fully isolated 24 GB instances.[3][5]

SpecificationWorkstation / Max-QServer Edition
ArchitectureBlackwell (GB202)Blackwell (GB202)
CUDA cores24,06424,064
Tensor Cores752 (5th gen)752 (5th gen)
RT Cores188 (4th gen)188 (4th gen)
Memory96 GB GDDR7 with ECC96 GB GDDR7 with ECC
Memory interface512-bit512-bit
Memory bandwidth1,792 GB/s1,597 GB/s
FP4 AI performance4,000 AI TOPS (with sparsity)4 PFLOPS
FP32 performance125 TFLOPS120 TFLOPS
RT Core performance380 TFLOPS355 TFLOPS
Media engines4x NVENC (9th gen), 4x NVDEC (6th gen)4x NVENC (9th gen), 4x NVDEC (6th gen)
InterfacePCIe Gen 5 x16PCIe Gen 5 x16
Display outputs4x DisplayPort 2.14x DisplayPort 2.1
MIGNot applicableUp to 4 instances of 24 GB
Max power600 W (WS) / 300 W (Max-Q)400 W to 600 W (configurable)

The Server Edition runs its memory at a slightly lower effective bandwidth (1,597 GB/s versus 1,792 GB/s on the Workstation Edition) to suit data-center thermal and power constraints. The ninth-generation NVENC engines add 4:2:2 H.264 and HEVC encoding support, and the sixth-generation NVDEC engines provide up to double the H.264 decoding throughput of the prior generation. The Workstation Edition launched at an approximate price of 8,565 US dollars.[5][6]

How does it compare to the RTX 6000 Ada Generation?

The RTX PRO 6000 Blackwell succeeds the RTX 6000 Ada Generation, which launched in late 2022 on the Ada Lovelace architecture using the AD102 die. The naming convention dropped the "Quadro" branding entirely and introduced the "RTX PRO" family name. The generational gains are substantial across every axis.

SpecificationRTX 6000 Ada GenerationRTX PRO 6000 Blackwell (Workstation)
ArchitectureAda Lovelace (AD102)Blackwell (GB202)
CUDA cores18,17624,064
Tensor Cores568 (4th gen)752 (5th gen)
RT Cores142 (3rd gen)188 (4th gen)
Memory48 GB GDDR6 with ECC96 GB GDDR7 with ECC
Memory interface384-bit512-bit
Memory bandwidth960 GB/s1,792 GB/s
FP32 performance91 TFLOPS125 TFLOPS
Max power300 W600 W
InterfacePCIe Gen 4 x16PCIe Gen 5 x16

The doubling of memory capacity to 96 GB and the addition of FP4 precision position the Blackwell card for large-language-model workloads that the Ada-generation part could not host on a single GPU. The trade-off is a doubling of the maximum board power to 600 W on the Workstation Edition, alongside an upgrade from a 384-bit GDDR6 bus to a 512-bit GDDR7 bus that nearly doubles memory bandwidth from 960 GB/s to 1,792 GB/s.[5]

What is the RTX PRO 6000 Blackwell used for?

NVIDIA positions the RTX PRO 6000 Blackwell as a unified platform for AI and graphics workloads that previously required separate hardware. The 96 GB of memory allows a single card to hold large generative models, supporting on-premises AI inference and fine-tuning without offloading to external accelerators. The fifth-generation Tensor Cores and FP4 support target high-throughput inference for large language models and generative AI, an area otherwise served by data-center accelerators such as the L40S.

Beyond AI, the fourth-generation RT Cores and high memory capacity accelerate ray-traced and neural rendering for media and entertainment, real-time 3D design, architecture, engineering and construction (AEC) workflows, manufacturing prototyping, scientific and engineering simulation, and physical AI development. In the Server Edition, MIG partitioning and multi-GPU scaling extend these capabilities to shared enterprise and cloud environments, where NVIDIA promoted RTX PRO Servers as a path to running agentic AI alongside visualization on mainstream enterprise systems.[2][3]

What is the RTX PRO 6000D and why was it created?

United States export controls progressively restricted the sale of NVIDIA's most capable accelerators to China, beginning with the A100 and H100 class and continuing through China-specific parts such as the H20. To remain in the Chinese market under these rules, NVIDIA developed a modified version of the RTX PRO 6000 reported as the RTX PRO 6000D, associated in reporting with the internal codename "B40."

According to reporting in mid-2025, the China variant retains the Blackwell architecture and GDDR7 memory but constrains memory bandwidth to roughly 1,398 GB/s, kept deliberately below the United States threshold of about 1.4 TB/s that governs which chips may be exported. Early reports, citing supply-chain outlet DigiTimes, described a card delivering around 1,100 GB/s of bidirectional bandwidth, fabricated on TSMC's N4 process, with shipments targeted for the third quarter of 2025. The RTX PRO 6000D notably lacks NVLink, relying on PCIe or external network interface cards such as ConnectX for multi-GPU communication, which sharply limits its usefulness for scaling large-model inference across many GPUs. Its reported retail price in China was around 50,000 yuan (roughly 7,000 US dollars).[7][8]

The variant met a difficult reception. In September 2025, reporting by the Financial Times and others, including Bloomberg, indicated that the Cyberspace Administration of China directed major technology companies such as Alibaba and ByteDance to halt testing and cancel orders for the RTX PRO 6000D. The directive was described as stronger than earlier guidance that had targeted the H20, and it pushed Chinese buyers toward domestic alternatives from suppliers such as Huawei, Cambricon, and Biren. Industry coverage noted that the chip's value proposition was weak relative to grey-market consumer cards: banned GeForce RTX 5090 boards reportedly circulated for around 3,500 US dollars while delivering competitive inference performance, undercutting the more expensive, bandwidth-limited 6000D.[9][10]

Why is the RTX PRO 6000 Blackwell significant?

The RTX PRO 6000 Blackwell marks the point at which NVIDIA's professional desktop line absorbed data-center-class memory capacity, with 96 GB on a single workstation card enabling local development and serving of large AI models. By spanning workstation, Max-Q, and server editions from one silicon design, it lets organizations standardize on a single GPU across desktops, dense multi-GPU workstations, and enterprise servers. At the same time, the RTX PRO 6000D episode illustrates how export controls reshape product strategy: a deliberately bandwidth-limited derivative, stripped of NVLink to satisfy regulatory thresholds, found little demand in a market that was simultaneously being steered by its own government toward domestic accelerators.

References

  1. RTX PRO 6000 Blackwell Series, NVIDIA
  2. NVIDIA Blackwell RTX PRO Comes to Workstations and Servers, NVIDIA Newsroom
  3. RTX PRO 6000 Blackwell Server Edition, NVIDIA
  4. RTX PRO 6000 Blackwell Max-Q Workstation Edition, NVIDIA
  5. RTX PRO 6000 Blackwell Workstation Edition, NVIDIA
  6. NVIDIA RTX PRO 6000 Blackwell Pricing, Thunder Compute
  7. Nvidia reportedly preparing RTX 6000D for Chinese market, Tom's Hardware
  8. Exploring the RTX Pro 6000D, Nvidia's China-only GPU, Tom's Hardware
  9. China Tells Companies to Stop Buying Nvidia's Repurposed AI Chip, Bloomberg
  10. China Reportedly Blocks NVIDIA RTX Pro 6000D, TrendForce

Improve this article

Add missing citations, update stale details, or suggest a clearer explanation. Every suggestion is reviewed for sourcing before it goes live.

1 revision by 1 contributors · full history

Suggest edit