Daily Tech Briefing
AI 科技速览
每天 5 分钟内学习 AI。获取最新的人工智能新闻,理解其重要性,并学习如何将其应用于您的工作。
NVIDIA Blog · 2026/7/21 15:36:43
NVIDIA Vera Rubin Driving Performance Per Watt, Lowest Token Cost for Partners Worldwide
AI 中文解读
NVIDIA Vera Rubin平台来了!这次核心亮点是它能在消耗同样电力的前提下,让AI运算速度提升10倍,大幅降低每个AI任务的计算成本。通俗来说,就像把一台原本又笨重又费电的老式服务器,升级成了更小、更省电、计算能力却翻了好几倍的微型工作站。这个平台不是简单拼凑现成零件,而是从芯片到整个机架系统都做了深度定制,甚至实现了机柜内不用一根线缆和风扇,组装时间从几小时缩短到一分钟。对于普通人来说,最直接的影响是未来用AI聊天、做图、处理文档时会更快更便宜。比如你用某个AI助手,背后的数据中心因为用了Vera Rubin,响应速度会提升,而服务商成本下降后,可能推出更低价甚至免费的高阶功能。同时,它支持更高温度的液冷散热,建数据中心的能耗和用水量都会减少,这对环保也是好消息。总之,这项技术让AI变得更“省钱”和“高效”,最终受益的是你我这样的普通用户。
NVIDIA Vera Rubin is here, and it’s going gigascale.
Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud, Microsoft Azure and Oracle Cloud Infrastructure. Spanning 350+ factory sites in 30 countries, Vera Rubin has the largest, most mature rack-scale supply chain ever assembled to meet customer compute demand.
The Vera Rubin platform is built from chip to grid to deliver the highest performance per watt and the lowest token cost. CoreWeave’s first benchmark on DeepSeek-R1 says it all: 10x more throughput per megawatt than Grace Blackwell NVL72 — landing directly on the metric that matters most for power-constrained AI factories.
Advancing Performance With Extreme Codesign
What makes this possible is extreme codesign across seven chips and five rack trays — Vera Rubin NVL72, Vera CPU rack, Groq 3 LPX, Spectrum-6 SPX and Vera BlueField-4 STX — all engineered as a single system rather than assembled from separate off-the-shelf products.
The NVIDIA Vera CPU is at its center. It redefines what an AI factory CPU can be. Designed and built for the agent era, its custom Olympus core delivers 2x single-threaded performance, 3x core-to-core bandwidth and 40% lower memory latency versus competing chiplet designs, making it the most efficient single-threaded CPU for the agentic workloads that matter most.
Accelerating AI Factories With Purpose-Built Networking
For networking, the platform’s sixth-generation NVLink scale-up delivers more than 2x throughput on complex workloads, 3x lower latency and 10x higher packet rates than off-the-shelf Ethernet. For scale-out, Spectrum-X Ethernet combines 102.4T Spectrum-6 switch systems, 1.6T ConnectX-9 SuperNICs, adaptive routing, advanced congestion control, telemetry and open operating system support, enabling 1.6x higher RDMA bandwidth than off-the-shelf Ethernet.
The world’s leading AI infrastructure builders — including CoreWeave, Microsoft, SpaceXAI and Tesla — are among the first to bring in Spectrum-6 switches to accelerate their AI factories. NVIDIA Photonics with co-packaged optics for scale-out — the industry’s first such switch in volume manufacturing — adds 5x lower power and 10x higher MTBI versus pluggable transceivers, with CoreWeave, Lambda and OCI among the first adopters.
Spectrum-XGS Ethernet extends performance across sites with 1.9x multi-site throughput because gigascale AI isn’t a single building problem.
And NVLink Fusion opens the NVIDIA infrastructure platform to third-party XPUs, giving partners a faster path to market on the proven NVLink scale-up stack and ecosystem.
Saving Setup Time, Water
NVIDIA’s three generations of rack-scale codesign produced a Vera Rubin NVL72 system with no cables, fans or hoses in the tray, cutting compute tray assembly time from hours to one minute.
A 45-degree Celsius liquid cooling inlet temperature design enables chiller-free dry-cooler operation. For new AI factories, this higher-temperature dry cooling along with the closed-loop liquid cooling system saves millions of gallons of water per megawatt annually.
Tuesday, July 21, 8:00 a.m. PT
NVIDIA Vera Rubin Powers Europe’s Open Model Era
Vera Rubin is delivering next-generation performance to Europe’s AI infrastructure.
It’s the foundation for a newly expanded Microsoft and Mistral partnership that brings frontier AI to the region, combining open European models with cloud and customer-controlled environments so governments and regulated industries can adopt it on their own terms.
Underpinning the partnership is a new multibillion-dollar agreement focused on expanding AI infrastructure in Europe. Mistral is adding its GPU capacity, drawing on thousands of the latest NVIDIA Vera Rubin GPUs to increase AI compute availability for customers and provide a shared platform for training, inference and large-scale deployment.
NVIDIA Vera Rubin is ramping into full production. The rack-scale AI supercomputer unifies seven new chips codesigne
分享
阅读原文 ↗