黄仁勋是否指10万块GPU服务器
⚠️ The most important question hardly anyone is asking is whether Jensen meant 100K+ GPUs, or 100k+ ...
Gary Marcus解读黄仁勋言论,分析Astra训练成本与性能提升的矛盾
Gary Marcus质疑黄仁勋是否指10万块GPU或10万块NVLink72服务器机架。每台NVLink72服务器包含72块Blackwell GPU。若指后者,训练Astra的硬件成本约为2500亿美元。EpochAI研究表明性能提升未超出趋势。行业面临训练成本指数增长与性能提升有限的矛盾。
⚠️ The most important question hardly anyone is asking is whether Jensen meant 100K+ GPUs, or 100k+ ...
⚠️ The most important question hardly anyone is asking is whether Jensen meant 100K+ GPUs, or 100k+ NVLink72 server racks (which contain 72 Blackwell GPUs). ⚠️ From his wording, it sure looks like the latter to me. If he indeed meant 100k+ NVLink72 server racks, we can infer that the hardware to train Astra sells for something like a quarter trillion dollars. Rental prices would perhaps be in the low tens of billions. For an improvement that @EpochAIResearch shows is not off trend. One key foundational problem with this whole industry (aside from technical limits of LLMS) is that you have two trends; exponential increases in training costs, modest increases in performance. Couple that with price wars and all of this is absolutely insane. It’d be like a gas company paying exponentially more money for each extra million barrels in the midst of a massive price war. That can’t last. Nor can this. Jensen Huang @JensenHuang @ChaseLochmiller @OpenAI GPT-6 Astra, trained on ~100K+ NVIDIA Grace Blackwell NVLink72. From ChatGPT to o1 to Astra in 4 years. AGI has arrived. Congratulations @OpenAI team. 400K GPUs coming online next. 🔗 View Quoted Tweet 💬 4 🔄 5 ❤️ 12 👀 2811 📊 5 ⚡