What we know
Amazon Web Services (AWS) and NVIDIA have announced plans to expand their collaboration by deploying an additional 2 million NVIDIA GPUs. AWS is expected to be the first cloud provider to offer NVIDIA’s next-generation GH200 Grace Hopper Superchips and the NVLink Switch System at scale. This expansion aims to support growing global demand for advanced AI computing, specifically targeting agentic AI workloads—autonomous agents—and physical AI applications such as robotics and simulation.
Why it matters
This announcement signals a significant increase in AI infrastructure capacity from two major industry players, reflecting the broader trend of hyperscale cloud providers investing heavily in AI hardware. AWS and NVIDIA’s longstanding partnership has been central to delivering high-performance computing for AI and machine learning workloads. The planned deployment of millions of GPUs and introduction of next-generation chips and networking technology could accelerate development and deployment of complex AI models and autonomous systems. The language used in source materials, such as describing AI as “smarter” or “stronger,” reflects vendor framing and is not independently validated.
What is still unknown
No independent tests of accuracy or energy savings yet. Key details such as the exact timeline for deployment, technical specifications, performance benchmarks, and potential customer impact remain unknown.
