Interestana
Home/News/Nvidia Unveils Blackwell GPU For AI Supercomputing
The Atlantic4 min read

By Interestana AI Editorial — AI-drafted, human-overseen. How we report

Nvidia Unveils Blackwell GPU For AI Supercomputing

Nvidia Unveils Blackwell GPU For AI Supercomputing

Nvidia announced its next-generation Blackwell GPU architecture on March 18, 2024, a significant advancement aimed at powering the next wave of artificial intelligence and data analytics. The Blackwell platform is engineered to handle the immense computational demands of training and deploying massive AI models, including large language models and generative AI applications. This new architecture succeeds the Hopper architecture, which has been a cornerstone of AI development for major cloud providers and research institutions.

The Blackwell GPU features a new chip design that integrates 208 billion transistors, a substantial increase over its predecessors. This enhanced density allows for greater processing power and efficiency. A key innovation is the second-generation Transformer Engine, which Nvidia states can accelerate deep learning training and inference by up to four times compared to the previous generation. This engine is crucial for optimizing the performance of complex neural networks that underpin modern AI.

Nvidia highlighted the Blackwell GPU's ability to deliver up to 1.4x higher performance for training and up to 30x faster inference for large language models. This performance leap is attributed to several architectural improvements, including a new unified memory system that provides 1.75 times more memory capacity at 1.5x the bandwidth compared to Hopper. The platform also introduces NVLink, Nvidia's high-speed interconnect technology, which has been upgraded to offer 1.8 terabytes per second of bidirectional bandwidth, enabling multiple GPUs to function as a single, massive processing unit.

The Blackwell architecture is designed for scalability, supporting configurations with up to 576 Blackwell GPUs connected via NVLink. This capability is essential for training the largest AI models, which can require immense computational resources. Nvidia also introduced the GB200 Superchip, which combines two Blackwell GPUs with a Grace CPU, offering a significant boost in performance and power efficiency for AI workloads. The GB200 is claimed to deliver up to 30 times the performance of the previous generation for large language model inference while consuming 25 times less energy.

Nvidia's announcement positions Blackwell as a critical infrastructure component for the burgeoning AI industry. Major cloud service providers, including Amazon Web Services (AWS), Google Cloud, and Microsoft Azure, have announced plans to adopt Blackwell GPUs for their AI services. Other prominent adopters include Meta Platforms and Oracle Cloud Infrastructure. The company also unveiled the DGX-700, an AI supercomputer built with 16 Blackwell GPUs, designed for enterprises to accelerate their AI initiatives. The focus on energy efficiency is also a significant aspect, with Nvidia emphasizing that Blackwell can reduce the total cost of ownership and carbon footprint for AI data centers.

Original source — read the full reporting at the publisher:

Read on The Atlantic

Get the weekly AI digest

AI news + new model releases, weekly. Drafted by our agents, reviewed by humans.

Read next