AmazonOfficial Platform UpdateWednesday, August 26, 20264 min read

AWS and NVIDIA expand partnership for next-gen AI infrastructure

About Amazon14h agoamazon
AWS and NVIDIA expand partnership for next-gen AI infrastructure
Executive Summary

AWS and NVIDIA will deploy 2 million additional GPUs and deepen collaboration across CPUs, networking, and robotics.

Source Lens

Official Platform Update

Direct platform communication. Highest-value for policy, product, and operational changes.

Impact Level

medium

Use this briefing to decide whether your team needs an immediate workflow, policy, or reporting change.

Key Stat / Trigger

No single quantitative trigger surfaced in this report.

Focus on the operational implication, not just the headline.

Relevant For
Brand SellersAgencies

Full Coverage

Key takeaways AWS will deploy 2 million additional NVIDIA GPUs across its global infrastructure in 2027–2028. NVIDIA Vera CPUs are coming to AWS, providing an additional compute option for agentic AI. The companies will build AI factories for the U. S. government, including 100,000 GPUs on secure AWS infrastructure.

Amazon Web Services (AWS) and NVIDIA announced a major expansion of their strategic collaboration.

Building on 16 years of joint innovation, the companies plan to deploy 2 million additional NVIDIA GPUs across AWS’s global infrastructure and deepen their work together across AI factories, CPUs, networking, open models, data processing, and robotics—delivering co-engineered solutions that help customers accelerate AI development at unprecedented scale.

The expansion complements Amazon’s own custom silicon, giving customers the freedom to choose the best compute for their specific workloads—whether that’s NVIDIA GPUs, AWS Trainium chips, or both working together. AWS and NVIDIA have worked together to bring cutting-edge AI capabilities to customers around the world for nearly two decades.

In fact, the two companies together launched the world’s first GPU-accelerated cloud instance on AWS, and today AWS offers the widest range of NVIDIA GPU solutions for customers. Now, as demand for AI accelerates, the two companies are taking that collaboration to a new level.

AI workloads are scaling at a rapid pace—from how models are trained and deployed, to how data is processed and used to power intelligent applications. Customers are moving from pilot to production across agentic AI, scientific discovery, enterprise automation, and robotics, and they need infrastructure that can keep pace.

“Customers want the freedom to choose the best tools for their AI workloads, and they want confidence that everything works seamlessly together,” said Matt Garman, CEO of AWS.

“That’s why we’ve invested deeply with NVIDIA to make AWS the best place to run NVIDIA AI technologies, optimizing performance across our infrastructure from networking and security to deployment. This expanded collaboration gives frontier labs, enterprises, and governments even more ways to build and deploy AI on AWS.”

“NVIDIA and AWS have built one of the great growth engines of the AI era, and demand is running ahead of every forecast,” said Jensen Huang, founder and CEO of NVIDIA. “For 16 years, we have scaled NVIDIA computing in the cloud together.

Now we are expanding our partnership across the full stack—GPUs, CPUs, networking, open models and software—to make agentic and physical AI real at an unprecedented pace and scale that only AWS and NVIDIA can deliver. This expansion reflects customers’ demand for NVIDIA’s platform on AWS.”

Scaling AI compute capacity At NVIDIA GTC 2026, AWS announced plans to add more than 1 million NVIDIA GPUs starting in 2026. Since then, demand has exceeded those expectations. AWS now plans to deploy an additional 2 million NVIDIA GPUs in 2027–2028 across its global infrastructure, including AI factories.

This capacity will power customer workloads ranging from agentic AI and scientific discovery to enterprise automation and physical AI. Customers are already seeing results from the collaboration—from faster drug discovery to more efficient fraud detection.

AWS and NVIDIA are also collaborating on advanced networking technology to connect GPUs more efficiently for large-scale AI training.

Bringing NVIDIA Vera CPUs to AWS AWS and NVIDIA are working to bring Vera CPU-based infrastructure to AWS, providing an additional option for agentic AI workloads that require high-performance CPU compute alongside accelerated infrastructure.

This will provide yet another option to configure AI infrastructure—consistent with AWS’s approach of offering the broadest possible set of compute choices rather than a one-size-fits-all solution.

Vera is built for the CPU work behind agentic AI and reinforcement learning, including code execution, tool use, sandboxing, analytics, data pipelines, and orchestration. As both a host CPU for accelerated systems and a standalone CPU for AI factory workloads, Vera keeps GPUs fed, agents responsive, and training loops moving.

Connecting AWS Trainium chips with NVLink Fusion At re:Invent 2025, AWS announced support for NVIDIA NVLink Fusion high-speed chip interconnect technology in next-generation Trainium chips.

Amazon’s Annapurna Labs will now work with NVIDIA’s new custom high-bandwidth memory (NVHBM) technology, which in partnership with memory suppliers, would give Trainium access to faster, more power-efficient memory.

Combined, NVHBM and NVLink Fusion make it possible for Annapurna Labs to tap NVIDIA’s custom memory technology and scale-up architecture to enhance performance and efficiency for AI workloads while seamlessly integrating Trainium and GPUs within a common rack-scale architecture.

Powering federal AI with the highest security Government agencies need secure AI infrastructu

Original Source

This briefing is based on reporting from About Amazon. Use the original post for full primary-source context.

View original
LinkedIn Post Generator

Style

Audience