Amazon just tripled its order of Nvidia chips over ‘surging demand’
Amazon just tripled its order of Nvidia chips over ‘surging demand’
亚马逊因“需求激增”将英伟达芯片订单量增加至三倍
Amazon and Nvidia just got a lot closer. The two companies announced Wednesday an expanded partnership that includes a deal to add another 2 million Nvidia GPU chips to Amazon’s data centers. These GPUs, which are designed to handle the heavy compute demands of training and running AI models, include Nvidia Blackwell Ultra, Rubin, and Rubin Ultra GPUs. The chips will head to Amazon Web Services’ data centers in 2027 and 2028.
亚马逊与英伟达的关系变得更加紧密。两家公司周三宣布扩大合作伙伴关系,其中包括一项协议:亚马逊将在其数据中心额外增加 200 万颗英伟达 GPU 芯片。这些专为处理训练和运行 AI 模型的高强度计算需求而设计的 GPU,包括英伟达的 Blackwell Ultra、Rubin 和 Rubin Ultra 系列。这些芯片将于 2027 年和 2028 年进入亚马逊云科技(AWS)的数据中心。
The announcement, made during Nvidia’s quarterly earnings call, comes just five months after Amazon agreed to deploy more than 1 million Nvidia GPUs across AWS infrastructure starting this year. Nvidia said in a statement that since then, “demand has exceeded those expectations.” Neither company shared financial terms. It’s unclear what the exact return will be for Nvidia. But considering GPU unit costs, the deal is worth tens of billions of dollars.
这一消息是在英伟达的季度财报电话会议上宣布的。就在五个月前,亚马逊刚同意从今年开始在 AWS 基础设施中部署超过 100 万颗英伟达 GPU。英伟达在一份声明中表示,自那时起,“需求已超出预期”。双方均未透露财务条款,目前尚不清楚英伟达的具体收益,但考虑到 GPU 的单价,这笔交易价值数百亿美元。
The announcement is notable not just for its size and the speed in which it grew, but also because it extends beyond Amazon buying more Nvidia chips. And it’s happening even as Amazon invests in its own potentially competing AI chips. Nvidia said Wednesday that its technology, including the networking hardware that connects thousands of GPUs into one system, as well as its open models, CPUs, data processing software, and robotics platform, will also be integrated across AWS. The companies said “surging demand” from startups, enterprises, AI labs, and even governments influenced the decision to work more closely.
这一公告不仅因其规模和增长速度引人注目,还因为它不仅仅局限于亚马逊购买更多的英伟达芯片。值得注意的是,尽管亚马逊也在投资其自身可能构成竞争的 AI 芯片,但双方的合作仍在深化。英伟达周三表示,其技术(包括将数千个 GPU 连接成一个系统的网络硬件,以及其开放模型、CPU、数据处理软件和机器人平台)也将整合到 AWS 中。两家公司表示,来自初创企业、大型企业、AI 实验室甚至政府的“需求激增”促成了双方更紧密的合作。
The expanded partnership comes as Amazon ramps up its own AI chip efforts — particularly with CPUs, which are the general purpose processors at the heart of servers. Amazon has been building its own chips to lessen its dependence on Nvidia and even compete with the chip giant. Amazon’s AI chief Peter DeSantis has said that AWS is in talks to sell its Trainium chips — which are a direct alternative to Nvidia’s H100 or Blackwell chips for deep learning workloads — to other companies for use in data centers. Amazon’s Arm-built Graviton CPU is also seen as a challenger to traditional server chips from Intel and AMD.
此次合作关系的扩大正值亚马逊加大自身 AI 芯片研发力度之际,尤其是作为服务器核心的通用处理器 CPU。亚马逊一直在开发自己的芯片,以减少对英伟达的依赖,甚至与这家芯片巨头展开竞争。亚马逊 AI 负责人 Peter DeSantis 表示,AWS 正在洽谈将其 Trainium 芯片(作为深度学习工作负载中英伟达 H100 或 Blackwell 芯片的直接替代品)出售给其他公司用于数据中心。亚马逊基于 Arm 架构开发的 Graviton CPU 也被视为英特尔和 AMD 传统服务器芯片的挑战者。
Amazon has said its custom chip business is growing, noting on its last earnings call that it crossed a $25 billion annualized revenue run rate, driven by $225 billion in total commitments from AI labs like Anthropic and OpenAI. But, it seems Nvidia is still the GOAT in the world of AI chips. With the 2 million GPU chips Amazon is adding to AWS starting in the third quarter, Nvidia also plans to send an unspecified number of Vera CPUs, “some integrated with Rubin, others standalone,” according to Nvidia CFO Colette Kress.
亚马逊表示其定制芯片业务正在增长,并在上一次财报电话会议上指出,得益于 Anthropic 和 OpenAI 等 AI 实验室总计 2250 亿美元的投入,其年化收入运行率已突破 250 亿美元。然而,英伟达似乎仍然是 AI 芯片领域的“史上最佳”(GOAT)。据英伟达首席财务官 Colette Kress 称,随着亚马逊从第三季度开始向 AWS 增加 200 万颗 GPU 芯片,英伟达还计划提供数量不详的 Vera CPU,“其中一些将与 Rubin 集成,另一些则作为独立产品”。
Nvidia CEO Jensen Huang has big plans for the company’s Vera CPUs, boasting back in May that he had found a “brand new $200 billion TAM” for the company. Aside from AWS, Kress said Wednesday that Nvidia expects Vera to be deployed by “every major hyperscaler, neocloud, AI lab, and system OEM, with shipments already underway to our lead partners,” which include Oracle and SpaceXAI.
英伟达首席执行官黄仁勋对公司的 Vera CPU 寄予厚望,他在五月份曾夸口称,他为公司发现了一个“价值 2000 亿美元的全新潜在市场(TAM)”。除了 AWS 之外,Kress 周三表示,英伟达预计 Vera 将被“每一家大型超大规模云服务商、新兴云服务商、AI 实验室和系统原始设备制造商(OEM)”部署,且“已开始向我们的主要合作伙伴发货”,其中包括甲骨文(Oracle)和 SpaceXAI。
The partnership is also extending to Amazon’s warehouse robots and enterprise offerings. Kress said Amazon plans to adopt Nvidia’s full physical AI stack to power its fleet of robots. The stack includes Omniverse (its simulation and digital twin platform); Cosmos (its world model platform); Isaac (its robotics development platform); and Jetson (computing hardware for robots and edge AI). This week, Nvidia also introduced a new version of Jetson designed as a more accessible robotics computer for “entry-level edge AI.”
此次合作还扩展到了亚马逊的仓库机器人和企业级产品。Kress 表示,亚马逊计划采用英伟达的全套物理 AI 技术栈来驱动其机器人车队。该技术栈包括 Omniverse(仿真和数字孪生平台)、Cosmos(世界模型平台)、Isaac(机器人开发平台)以及 Jetson(用于机器人和边缘 AI 的计算硬件)。本周,英伟达还推出了新版 Jetson,旨在作为一种更易于使用的机器人计算机,面向“入门级边缘 AI”。
On the enterprise side, AWS will serve Nvidia’s Nemotron family of open models on Amazon Bedrock, its managed foundation model platform, and SageMaker, its managed cloud service.
在企业端,AWS 将在其托管基础模型平台 Amazon Bedrock 和托管云服务 SageMaker 上提供英伟达的 Nemotron 系列开放模型。
Nvidia also reported Wednesday that it recorded sales of $96.2 billion for the second quarter, beating analyst estimates. Data center revenue made up the majority of Nvidia’s sales for the quarter at $89 billion, up 117% from a year ago. Nvidia said it expects revenue to reach $108 billion in the third quarter, some of which will come from its next-gen Rubin GPUs. Nvidia said it began production shipments this quarter. Investors have been looking out for Rubin’s initial Q3 sales for signs that demand will continue into Nvidia’s next generation of hardware.
英伟达周三还公布了第二季度财报,销售额达到 962 亿美元,超过了分析师的预期。数据中心业务贡献了该季度销售额的大部分,达到 890 亿美元,同比增长 117%。英伟达预计第三季度营收将达到 1080 亿美元,其中部分将来自其下一代 Rubin GPU。英伟达表示,本季度已开始进行生产发货。投资者一直在关注 Rubin 在第三季度的初步销售情况,以寻找需求将持续到英伟达下一代硬件的迹象。
Nvidia has committed $279 billion to secure supply and manufacturing capacity for current and future data-center projects, up substantially from $119 billion last quarter, as the chipmaker looks to secure memory and manufacturing capacity to meet AI demand over the next few years. That commitment includes $92 billion in projected spending for the rest of the fiscal year and another $87 billion in fiscal year 2028.
英伟达已投入 2790 亿美元用于确保当前和未来数据中心项目的供应和制造能力,较上一季度的 1190 亿美元大幅增加。这家芯片制造商正寻求锁定内存和制造产能,以满足未来几年的 AI 需求。这一承诺包括本财年剩余时间预计支出的 920 亿美元,以及 2028 财年的另外 870 亿美元。
“The thing that matters for the industry is that AI is now doing productive and useful work,” Huang said during Wednesday’s call. “AI is generating profitable tokens… If we had more compute, we could generate more profitable tokens, which results in more profit for all of the services. This is the exact phase where we’re at, which is the reason why everybody’s leaning in.” Investors will be watching to see if additional compute indeed translates so neatly into additional profits as AI companies pour hundreds of billions of dollars into infrastructure.
“对行业而言,重要的是 AI 现在正在进行富有成效且有用的工作,”黄仁勋在周三的电话会议上说。“AI 正在生成可盈利的 Token……如果我们有更多的算力,就能生成更多可盈利的 Token,从而为所有服务带来更多利润。这正是我们所处的阶段,也是每个人都在全力投入的原因。”随着 AI 公司向基础设施投入数千亿美元,投资者将密切关注额外的算力是否真的能如此顺畅地转化为额外的利润。