行业多源确认

爆料称 DeepSeek 年投入约 10 亿美元,其中 70% 用于训练

精选理由

有人扒出 DeepSeek 一年只花 10 亿美元、七成砸训练,还拿游戏卡跑小模型推理,抠成本思路挺值得看看。

一条 X 帖子称 DeepSeek 的年算力投入约 10 亿美元,其中 70% 用在训练上。帖中还提到 DeepSeek 正使用 NVIDIA 游戏显卡(GeForce 系列 GPU)为较小的模型跑推理。作者认为这是合理的省钱做法,比如用 Jev 式小模型或定制版 Qwen 35B 处理日常任务,把大集群留给训练。该说法来自社交媒体爆料,DeepSeek 官方尚未证实。

原文 · Teortaxes

$1B makes sense. 70% on training is logical > DeepSeek is using NVIDIA gaming GPUs to run inference on smaller models this is news to me but tbh is an obvious move. Imagine how much work you could speed up just with Jev-style small models, or a custom-made Qwen 35B… I want it https://t.co/pYnS3C85lW