Vera Rubin has moved from announcement into industrial ramp, and customers are now committing to it in public. NVIDIA said on July 21 that Vera Rubin NVL72 production is ramping across more than 350 factory sites in 30 countries with racks running at four cloud providers, and Wccftech reported hardware engineering chief Andrew Bell saying about a dozen partners could produce up to 1,000 Vera racks a day, which is stated capacity rather than documented output. CoreWeave says it was the first cloud provider to bring up and validate the NVL72 platform, built from 72 Rubin GPUs, 36 Vera CPUs, ConnectX-9 SuperNICs, BlueField-4 DPUs and NVLink 6. On August 4 Elon Musk said SpaceX had committed to NVIDIA GPUs exclusively, with Wccftech reporting Vera Rubin as the platform for its AI data centers and company targets of 2 gigawatts of compute by year-end and 10 gigawatts by the end of 2027; SpaceX separately said it is working with NVIDIA on the compute payload for its Starmind AI1 satellite, using Rubin GPUs and Vera CPUs. Those capacity numbers are forward-looking targets, not disclosed orders or deployed capacity. The configuration is less settled than the ramp suggests: Wccftech reported on July 26, citing GF Securities, that NVIDIA may halve the CPU-side LPDDR5X memory planned for NVL72 systems, following a supply-driven SOCAMM reduction TrendForce reported in June, while NVIDIA's public specifications still list 54TB of CPU memory per rack and the 20.7TB of GPU-side HBM4 is unchanged in that scenario. Further out, CNBC reported that Kyber NVL144 has been pushed to 2028 because the specialized PCB midplane remains difficult to manufacture, with SemiAnalysis also saying NVL576 could be delayed or limited to small volumes and the NVL72x2 back-to-back architecture canceled; NVIDIA did not comment. For anyone sizing a 2027 or 2028 cluster, host-memory, rack-cost and delivery assumptions should be modeled as ranges: the $21 million Rubin Ultra rack figure and the roughly $1.5 million HBM4e component in it are BofA Global Research and Morgan Stanley estimates, not NVIDIA list prices.
The second shift is how much of NVIDIA's activity is now capital and supply commitment rather than product launch. It signed a $1.5 billion multi-year advanced packaging and test agreement with Amkor on July 23 that includes an NVIDIA prepayment supporting Amkor's U.S. packaging expansion; announced a long-term partnership with Safe Superintelligence on July 27 giving Ilya Sutskever's lab Vera Rubin access and a tenfold increase in compute, with an investment NVIDIA disclosed but did not size and Reuters reported at $5 billion citing a person briefed on the deal; joined NAVER and Brookfield on July 25 in a proposed expansion of the GAK Sejong AI factory from 55 MW to 200 MW by 2028, funded by up to $9 billion from Brookfield and a planned $1 billion NVIDIA investment; and signed letters of intent with SK Group on July 24 covering a $500-billion-plus initiative including an SK Telecom AI factory of up to 2 gigawatts on Vera Rubin DSX with a first phase targeted for 2027 and long-term HBM4 supply and co-development with SK hynix. Seoul valued the broader July 24 summit package at a reported $950 billion, a figure that combines multi-year cooperation agreements rather than completed purchases. Smaller checks are going out too, including a reported 237.9 crore rupee allocation in Sarvam AI's roughly $74 million Series B extension, following Sarvam's plan for a trillion-plus-parameter model on 10,000 Blackwell GPUs. On the software side, NVIDIA said on August 4 that it is open-sourcing the cuFile APIs and the storage stack beneath them, giving developers a contribution path to the technology that lets GPUs initiate storage reads and writes directly. Policy is the last thread: NVIDIA launched the Open Secure AI Alliance with more than 30 partners on July 27 and signed a July 24 letter with 24 other organizations urging U.S. policymakers to preserve access to open-weight models, which Jensen Huang promoted in his first post on X. Its China access is still decided elsewhere; Reuters-syndicated reporting on July 8 said Beijing was weighing limited H200 approvals for Alibaba, ByteDance and DeepSeek, potentially fewer than 200,000 chips. Investors have separately marked the stock down about 16% from its May 14 peak, removing roughly $1 trillion of market value and putting it near 18 times forward earnings, a rotation toward memory and storage names rather than reported evidence that GPU demand has weakened.