
NVIDIA CEO Says GPT-6 Astra Used 100,000 Blackwell GPUs

NVIDIA CEO Says GPT-6 Astra Used 100,000 Blackwell GPUs
WEEX View
- The key variable is whether NVIDIA’s disclosed cluster scale becomes a new benchmark for frontier model training rather than a one-off build. If it does, access to advanced GPUs and interconnect capacity may become even more concentrated among a small group of top AI companies.
- The report also points to a second watchpoint: generation gap matters as much as raw chip count. Blackwell, Hopper, and domestic alternatives are not directly interchangeable, so comparisons based only on total units may understate the performance gap.
- For crypto-adjacent AI narratives, the practical implication is not immediate token impact but renewed focus on compute scarcity, infrastructure bottlenecks, and whether smaller players can realistically compete without access to top-tier centralized hardware.
NVIDIA CEO Jensen Huang said GPT-6 Astra was trained using more than 100,000 Grace Blackwell GPUs connected in a high-speed NVLink72 cluster, according to the company disclosure cited in the report. Huang also said the next batch would use 400,000 GPUs, though the specific models were not identified.
The disclosure centers on the hardware used to train GPT-6 Astra. Huang said the system used more than 100,000 Grace Blackwell GPUs linked through NVLink72, NVIDIA’s architecture for placing 72 GPUs within the same high-speed interconnect domain. The report said the next training batch is planned at 400,000 GPUs, but did not specify whether those systems would also use Blackwell.
The report framed the disclosure against the current AI infrastructure gap between U.S. and Chinese model developers. ByteDance was described as the closest among major Chinese model companies, with about 36,000 B200 GPUs connected. Its domestic clusters were said to rely mainly on Hopper-based H20 and H800 systems.
Other capacity figures in the report remain less clear. Kimi was said to have obtained around 20,000 Hopper GPUs through Alibaba, specifically H200 units, although Alibaba denied that claim. DeepSeek has not disclosed the full training hardware used for V4, but leaked information cited in the report said the company had about 20,000 H-equivalent compute units in May, including roughly 16,000 units of Huawei 950 capacity.
The report also noted that the Blackwell platform used for Astra is no longer NVIDIA’s newest generation. Vera Rubin has already entered full-scale production, and NVIDIA estimates that training large mixture-of-experts models with Rubin could require only a quarter of the GPUs needed with Blackwell. No Chinese company has publicly disclosed using 100,000 advanced GPUs of the same generation to train a single model, according to the report.
Why It Matters
The disclosure matters because it sharpens the divide between frontier AI development and the broader market narrative around model competition. At the top end, performance is increasingly tied not just to algorithms or data, but to who can assemble, power, and interconnect extremely large clusters of the newest chips.
That has broader implications beyond AI vendors themselves. The more large-model progress depends on scarce, tightly integrated hardware stacks, the more strategic weight shifts toward semiconductor supply, networking architecture, and access to advanced compute. For crypto-linked AI sectors, that keeps the focus on infrastructure constraints rather than simple enthusiasm around AI branding.
This content is provided for general informational purposes only and doesn't constitute financial, investment, legal, or tax advice. Any events, rewards, online promotions, or related information mentioned herein should not be considered a recommendation, solicitation, or invitation to purchase, sell, trade, or otherwise deal in any crypto assets. Crypto assets are highly volatile and may result in loss. The availability of WEEX services, products, and related events may vary by region. You are responsible for ensuring that your participation is in accordance with applicable local laws and regulations.
About WEEX View
WEEX View is a crypto analysis and intelligence hub, covering the latest in Web3, AI, and global markets. Get independent research and in-depth insights to stay ahead of market trends and trading opportunities.
Latest articles
MoreEthereum Advances Draft Plan for Gas Payments Without Direct ETH
Ethereum approved EIP-8141 at an August 27 meeting, advancing a draft plan to let transaction fees be paid by another account or application, with proposed inclusion in the Hegota upgrade in 2027.
CFTC Seeks Dismissal of CME Suit Over Kalshi Bitcoin Perpetuals
The CFTC asked a Washington, D.C. court to dismiss CME's lawsuit challenging Kalshi's bitcoin perpetual futures product, setting up a key response deadline as the dispute tests how regulated crypto derivatives may be classified in the U.S.
Copper Hits Record on LME as Tariffs and Supply Tighten
Copper rose to a record $14,533 a ton on the London Metal Exchange, with U.S. tariff pressure, stronger demand and supply constraints driving the move as traders shift metal into the U.S.
US Senate Sets September 15 Vote on Clarity Law
The U.S. Senate is set to vote on the Clarity Law on September 15, with Senator Cynthia Lummis warning that failure in this Congress could push the next practical window for digital-asset legislation to 2030.




