Home Blogs Data Center Explorer AWS and Nvidia partner on Project Ceiba, a GPU-powered AI supercomputer

AWS and Nvidia partner on Project Ceiba, a GPU-powered AI supercomputer

News

Nov 30, 20233 mins

CPUs and ProcessorsData CenterGenerative AI

The companies are extending their AI partnership, and one key initiative is a supercomputer that will be integrated with AWS services and used by Nvidia’s own R&D teams.

shutterstock 324149159 cloud computing building blocks abstract sky with polygons and cumulus clouds

Credit: Shutterstock

Amazon Web Services and Nvidia have announced an expansion of their alliance that includes plans to add supercomputing capabilities to AWS’s artificial intelligence (AI) infrastructure. The companies announced the news at the AWS re:Invent conference in Las Vegas.

The biggest of the initiatives is Project Ceiba, a supercomputer that will be hosted by AWS for Nvidia’s own research and development teams. It will feature 16,384 Nvidia GH200 Superchips and be capable of processing 65 exaflops of AI, the companies said. The Project Ceiba supercomputer will be integrated with a number of AWS services, including Amazon Virtual Private Cloud (VPC) encrypted networking and Amazon Elastic Block Store high-performance block storage.

Nvidia plans to use the supercomputer for research and development to advance AI for LLMs, graphics and simulation, digital biology, robotics, self-driving cars, Earth-2 climate prediction and more.

New Amazon EC2 G6e instances featuring Nvidia L40S GPUs and G6 instances powered by L4 GPUs are also in the works, AWS announced. L4 GPUs are scaled back from the Hopper H100 but offer much more power efficiency. These new instances are aimed at startups, enterprises, and researchers looking to experiment with AI.

Nvidia also shared plans to integrate its NeMo Retriever microservice into AWS to help users with their development of generative AI tools like chatbots. NeMo Retriever is a generative AI microservice that enables enterprises to connect custom LLMs to enterprise data, so the company can generate proper AI responses based on their own data.

“Generative AI is transforming cloud workloads and putting accelerated computing at the foundation of diverse content generation,” said Jensen Huang, founder and CEO of Nvidia, in a statement. “Driven by a common mission to deliver cost-effective, state-of-the-art generative AI to every customer, Nvidia and AWS are collaborating across the entire computing stack, spanning AI infrastructure, acceleration libraries, foundation models, and generative AI services.”

In other news, AWS will be the first cloud provider to bring Nvidia’s GH200 Grace Hopper Superchips to the cloud. The Nvidia GH200 NVL32 multi-node platform connects 32 Grace Hopper Superchips by Nvidia’s NVLink and NVSwitch interconnects. The platform will be available on Amazon Elastic Compute Cloud (EC2) instances connected with Amazon’s networking, virtualization (AWS Nitro System), and hyper-scale clustering (Amazon EC2 UltraClusters).

AWS will host Nvidia’s DGX Cloud cluster of GPUs for AI. DGX Cloud on AWS will accelerate training of generative AI and LLMs that can reach beyond 1 trillion parameters.

by Andy Patrizio

Andy Patrizio is a freelance journalist based in southern California who has covered the computer industry for 20 years and has built every x86 PC he’s ever owned, laptops not included.

The opinions expressed in this blog are those of the author and do not necessarily represent those of ITworld, Network World, its parent, subsidiary or affiliated companies.

Americas

Topics

About

Policies

Our Network

More

AWS and Nvidia partner on Project Ceiba, a GPU-powered AI supercomputer

The companies are extending their AI partnership, and one key initiative is a supercomputer that will be integrated with AWS services and used by Nvidia’s own R&D teams.

More from this author

Everyone but Nvidia joins forces for new AI interconnect

Ampere updates roadmap, heads to 256 cores

AMD holds steady against Intel in Q1

Broadcom launches 400G Ethernet adapters

HPE updates block storage services

ZutaCore launches liquid cooling for advanced Nvidia chips

Nvidia to build supercomputer for federal AI research

High-bandwidth memory nearly sold out until 2026

Most popular authors

Show me more

Cisco patches actively exploited zero-day flaw in Nexus switches

Nokia to buy optical networker Infinera for $2.3 billion

French antitrust charges threaten Nvidia amid AI chip market surge

Has the hype around ‘Internet of Things’ paid off? | Ep. 145

Episode 1: Understanding Cisco’s Converged SDN Transport

Episode 2: Pluggable Optics and the Internet for the Future

How to use the stat command

The SL command easter egg

How to use the shuf command

AWS and Nvidia partner on Project Ceiba, a GPU-powered AI supercomputer

The companies are extending their AI partnership, and one key initiative is a supercomputer that will be integrated with AWS services and used by Nvidia’s own R&D teams.

Related content

Pure Storage adds AI features for security and performance

Nvidia teases next-generation Rubin platform, shares physical AI vision

Intel launches sixth-generation Xeon processor line

AMD updates Instinct data center GPU line

Newsletter Promo Module Test

More from this author

Everyone but Nvidia joins forces for new AI interconnect

Ampere updates roadmap, heads to 256 cores

AMD holds steady against Intel in Q1

Broadcom launches 400G Ethernet adapters

HPE updates block storage services

ZutaCore launches liquid cooling for advanced Nvidia chips

Nvidia to build supercomputer for federal AI research

High-bandwidth memory nearly sold out until 2026

Most popular authors

Show me more

Cisco patches actively exploited zero-day flaw in Nexus switches

Nokia to buy optical networker Infinera for $2.3 billion

French antitrust charges threaten Nvidia amid AI chip market surge

Has the hype around ‘Internet of Things’ paid off? | Ep. 145

Episode 1: Understanding Cisco’s Converged SDN Transport

Episode 2: Pluggable Optics and the Internet for the Future

How to use the stat command

The SL command easter egg

How to use the shuf command