China National AI Fund Commits $1B To DeepSeek With Voting Rights
Reactions from ranked influencers
3 postsmeanwhile: «The China National AI Industry Investment Fund invested in DeepSeek and received voting rights. The fund has committed RMB $1B.» I suspect Wenfeng can be a bit of a drama queen. Consider that «DeepSeek» is a quote from Li Sao. This state may prove more benevolent.
"The government won't give me a cent" https://twitter.com/teortaxesTex/status/2080158170857574747
Wenfeng's investor call should be read skeptically. It has several curious contradictions. He says he won't do chips "if possible" (has actually been working on chips for years). The part about the government not giving him a cent (got $150M investment). And… $3B tops for compute? At $25K/pop, that's 120K H200s, but *he can't take so much* – the sales cap is 75K/entity. Add 20K he has (and almost certainly those are heavily H800s, H20s, some H100s – garbage), it still isn't enough. And yet, his job listing "AI Computing Cluster Performance and Reliability Engineer" explicitly talks about a 100K cluster they need to keep operational, filled with "supernodes" among all else. "on your first day, you'll be managing a 100.000 card cluster". It is training-oriented; the listing is full of autism about whole-system stability (inference nodes can fail independently). There's a lot of concern about different "batches" of hardware, which also shouldn't be a big issue with H200s; it's a very well understood device from a mature production line. On the other hand, it says NCCL, NVLink, "Understanding of RDMA/InfiniBand/GPU architecture or performance tuning (a plus, but not necessary)", so I guess he's not lying that he intends to NVidia-maxx as much as possible (this also plays to DeepSeek's strengths, they have sublime mastery of Nvidia stack at this point, arguably more than Nvidia itself). And on the gripping hand, there are listings that mention Ascend and "CPU/GPU/NPU" ("High-performance operator/communication/compiler engineer", "Supercomputing Cluster R&D Engineer"). I think it's a tightly coupled heterogenous system that'll, for example, use Nvidia for weight updates and SuperPoDs for rollouts, in the way we can infer from V4 paper. As Zephyr notes, 950s are relatively less inefficient in inference, and the software is definitely way more mature https://x.com/zephyr_z9/status/2080368613064945803. Eventually Ascends will be supplemented or supplanted with in-house Whale Compute which is even more inference-optimal, and as the availability of Huawei hardware grows and Nvidia's dries up, Ascends will probably get shifted to training as well. In short, it's tight, but I think he may well have the equivalent of "200K 950s" by the end of 2026 Q4, and his people will be able to conduct serious research that leads to next-generation Whales. An OOM scaled-up pretraining will have to come later.
meanwhile: «The China National AI Industry Investment Fund invested in DeepSeek and received voting rights. The fund has committed RMB $1B.» I suspect Wenfeng can be a bit of a drama queen. Consider that «DeepSeek» is a quote from Li Sao. This state may prove more benevolent. https://twitter.com/zephyr_z9/status/2080158728171724926
Interesting
@teortaxesTex This is a speech on May 20, and the allocation of the National Fund was not handed over until June 17. Before that, it seemed that DeepSeek had not received any preferential treatment, otherwise it would not have been until April this year that there were only a few thousand H100
Combined views
18.2K
3 posts, first seen 8h ago