Gary Marcus

Gary Marcus

@garymarcus · Twitter ·

⚠️ The most important question hardly anyone is asking is whether Jensen meant 100K+ GPUs, or 100k+ NVLink72 server racks (which contain 72 Blackwell GPUs). ⚠️ From his wording, it sure looks like the latter to me. If he indeed meant 100k+ NVLink72 server racks, we can infer that the hardware to train Astra sells for something like a quarter trillion dollars. Rental prices would perhaps be in the low tens of billions. For an improvement that @EpochAIResearch shows is not off trend. One key foundational problem with this whole industry (aside from technical limits of LLMS) is that you have two trends; exponential increases in training costs, modest increases in performance. Couple that with price wars and all of this is absolutely insane. It’d be like a gas company paying exponentially more money for each extra million barrels in the midst of a massive price war. That can’t last. Nor can this.

Jensen Huang

Jensen Huang

@ChaseLochmiller @OpenAI GPT-6 Astra, trained on ~100K+ NVIDIA Grace Blackwell NVLink72. From ChatGPT to o1 to Astra in 4 years. AGI has arrived. Congratulations @OpenAI team. 400K GPUs coming online next.