Friday, September 11, 2026
AI 인프라 · 뉴스 & 분석
데이터센터리포트
데이터센터 · 리포트

AWS는 자체 인프라에 추가로 200만 개의 Nvidia GPU를 도입하기로 약속했다.

이 전례 없는 하이퍼스케일러의 용량 공약은 글로벌 GPU 공급망에 심각한 부담을 가중시키고, 랙 레벨 밀도 업그레이드를 가속화하며, 전력 및 냉각 네트워크의 급속한 확장을 강요할 것이다.
업계 전문지Slicast · August 28, 2026 · 글로벌 · 출처: Data Center Dynamics
중요도 95

Amazon Web Services (AWS) has committed to expanding its fleet of Nvidia GPUs, agreeing to deploy an additional two million units across its global infrastructure. The expanded agreement also deepens collaboration between the two companies on AI factories, CPUs, networking, open models, data processing, and robotics.

AWS will roll out the new GPUs through 2027 and 2028. As part of the initiative, the cloud provider will introduce Nvidia Vera CPU-based infrastructure to its platform, extend Nvidia NVLink Fusion with custom Nvidia high-bandwidth memory, and construct AI factories for the U.S. government featuring 100,000 GPUs on secure AWS infrastructure. Additionally, Nvidia’s Nemotron models will be made available through AWS, while Amazon Robotics will integrate Nvidia’s physical AI platform to advance joint robotics initiatives.

During Nvidia GTC 2026, AWS initially announced plans to bring online more than one million Nvidia GPUs to its platform in 2026, but demand has since outpaced that projection. The newly added two million GPUs will encompass Blackwell Ultra, Rubin, and Rubin Ultra architectures. AWS will also expand its Blackwell capacity by deploying Nvidia RTX Pro 4500 Blackwell Server Edition GPUs for its EC2 G7 instances, which offer 4.6x AI inference performance and 2.1x graphics performance compared to previous-generation G6 instances.

“Customers want the freedom to choose the best tools for their AI workloads, and they want confidence that everything works seamlessly together,” said Matt Garman, CEO of AWS. “That’s why we’ve invested deeply with Nvidia to make AWS the best place to run Nvidia AI technologies, optimizing performance across our infrastructure from networking and security to deployment. This expanded collaboration gives frontier labs, enterprises and governments even more ways to build and deploy AI on AWS.”

“Nvidia and AWS have built one of the great growth engines of the AI era, and demand is running ahead of every forecast,” said Jensen Huang, founder and CEO of Nvidia. “For 16 years, we have scaled Nvidia computing in the cloud together. Now, we are expanding our partnership across the full stack — GPUs, CPUs, networking, open models and software — to make agentic and physical AI real at an unprecedented pace and scale that only AWS and Nvidia can deliver. This expansion reflects customers’ demand for Nvidia's platform on AWS.”

While scaling its fleet, AWS continues to operate older hardware. CEO Matt Garman has previously noted that the company still runs six-year-old A100 servers, stating it has “never retired an A100 server.”

The announcement follows Nvidia’s Q2 2026 earnings report, released this week. The company posted $96.2 billion in revenue for the quarter, representing an 18 percent increase from the prior quarter and a 106 percent year-over-year jump. Both GAAP and non-GAAP gross margins held steady at 75 percent.

원문 보기