OpenAI Signs $38 Billion Cloud Deal with Amazon Web Services

Nov 4, 2025
OpenAI has entered a seven-year, $38 billion agreement with Amazon Web Services to power its AI workloads using AWS infrastructure, including hundreds of thousands of NVIDIA GPUs and the ability to scale to tens of millions of CPUs by 2026.
OpenAI Signs $38 Billion Cloud Deal with Amazon Web Services

OpenAI has signed a seven-year, $38 billion agreement with Amazon Web Services to run its advanced AI workloads, announced in a press release. The partnership gives OpenAI immediate access to AWS's infrastructure, including hundreds of thousands of NVIDIA GPUs, and the capability to scale to tens of millions of CPUs by 2026, with room to expand further in 2027 and beyond.

Under the deal, AWS will provide OpenAI with Amazon EC2 UltraServers designed for large-scale AI processing. These clusters connect NVIDIA GB200 and GB300 accelerators on a single network to support both inference for ChatGPT and training of next-generation models. AWS said the infrastructure is optimized for performance and reliability across OpenAI’s evolving needs.

OpenAI CEO Sam Altman said the partnership strengthens the compute ecosystem required to scale frontier AI systems. AWS CEO Matt Garman stated that the company’s infrastructure will serve as a backbone for OpenAI’s workloads. The companies have previously collaborated to make OpenAI’s models available through Amazon Bedrock.

The agreement follows OpenAI’s recent restructuring, which allows it to source cloud services beyond Microsoft Corporation, and aligns with its broader push to expand global computing capacity through partnerships with major cloud and chip providers.

We hope you enjoyed this article

Consider subscribing to one of our newsletters like Enterprise AI Brief, Silicon Brief or Daily AI Brief.

Also, consider following us on social media:

Free newsletter

Silicon Brief

Weekly coverage of AI hardware developments including chips, GPUs, cloud platforms, and data center technology.

Whitepaper

Tensordyne Napier: What If One Rack Could Do the Work of Nine?

Tensordyne

This Tensordyne whitepaper presents Napier, an inference-focused AI processor and rack-scale system based on the company’s TDN Math logarithmic number system. It examines infrastructure requirements for large mixture-of-experts and agentic models, compares major inference architecture approaches, and details the TDN AIP processor, TDN72 pod, TDN Link fabric, and Napier Ultra configuration. The paper reports simulation-based performance, cost, and accuracy-validation results, including Tensordyne’s projected comparison of one Napier rack with a nine-rack Nvidia Rubin plus Groq deployment; the chip is reported as taped out and in fabrication.

Read more
Free, six days a week

Daily AI Brief: the AI news that matters, in your inbox.