What is DeepSeek?
DeepSeek is a Chinese artificial intelligence startup based in Hangzhou, founded in 2023 by Liang Wenfeng, known for developing open-source AI models.
Company timeline
- Apr 2025 – Brad Gerstner said that if the US bans DeepSeek, more people would work with Huawei chips and DeepSeek to optimize it, and other countries would use it.
- Apr 2026 – Dylan Patel stated that DeepSeek’s cost on GPT-4 was 1/600 the cost.
- Apr 2026 – Reiner Pope said the DeepSeek paper reports they do a lot of expert parallelism.
- May 2026 – Kyle Corbitt explained that GRPO took off because DeepSeek did engineering work around scaling it and released an actual artifact model.
- Jun 2026 – Carina Hong reported that the best DeepSeek LM scored 103 out of 120 on an exam, while the best human scored 110.
- Jul 2026 – Jason Calacanis said that models from Nvidia and Google’s Gemma family aren’t sufficient to replace what we have from DeepSeek and other Chinese models.
- Jul 2026 – Harry Stebbings reported that the Chinese government may deny access to overseas users for some Chinese open source models.
- Jul 2026 – Jay V observed that DeepSeek’s usage dipped around the release of GLM but then bounced back.
- Jul 2026 – Aman Sanger noted that if a user is just loading DeepSeek before the first inference, they might wait about 30 minutes.
- Aug 2026 – Nathaniel Whittemore said that Wall Street analysts and others are finally beginning to update their priors and recognize that Chinese labs aren’t, in most cases, selling frontier intelligence for pennies on the dollar the way they had previously assumed.
Where it appears in the record
Every line below is attributed to a named speaker.
Axiom Math scored 120/120 on the 2025 Putnam exam, beating the best LLM (DeepSeek at 103/120) and the best human contestant (110/120).
“The best LM deepseek got 103 points out of a 120 point exam. The best human obviously we now know is a student from either MIT or Chicago. We don't know which one because they don't announce the top five winner score got 110 and we got 120.”Carina Hong · 3 Jun 2026
GRPO became dominant not due to algorithmic novelty but because DeepSeek did the scaling engineering and shipped a working artifact model that validated it.
“I think the reason GRPO specifically like that algorithm and that acronym like, you know, very concretely took off was not necessarily because it was like a big quantum leap on what came before. It was because DeepSeek did a lot of engineering work around actually scaling it and released an actual artifact model that worked really well with it.”Kyle Corbitt · 1 May 2026
DeepSeek delivered GPT-4-level performance at 1/600th the inference cost of GPT-4.
“Deep Seek, for example, on GPT-4 was 1/600 the cost.”Dylan Patel · 23 Apr 2026
DeepSeek demonstrates that natural language verification with meta-verification is sufficient for mathematical reasoning at scale, without requiring formal proof systems like Lean.
“It's interesting that natural language verification with some sort of meta-verification seems to work so far in the published literature.”Grant Sanderson · 30 Jun 2026
DeepSeek uses expert parallelism rather than pipeline parallelism for frontier-scale inference, a deployment choice worth tracking as MoE models proliferate.
“The DeepSeek paper reports what they do, which is that they just do a lot of expert parallelism.”Reiner Pope · 29 Apr 2026
DeepSeek raised approximately 50 billion RMB (roughly $7.5 billion USD) from Tencent, CATL, and state funds including China's national AI industry investment fund.
“It has raised about 50 billion grand which is like 7.5 USD investors include 10 cent and a few other and oh and also CL the battery maker and some state funds including the national artificial intelligence industry investment fund.”Jordan Schneider · 24 Jul 2026
Nvidia Nemotron Ultra uses Mamba (hybrid) attention, positioning it architecturally differently from Chinese labs such as DeepSeek and MiMo that use sparse attention.
“In the Nvidia Ne Neumitron Ultra used the Mamba attention, whereas, you know, we see, you know, Deep Seek sparse attention and then the MiMo MSA, whatever that stands for, MiMo sparse attention.”Finbarr Timbers · 16 Jun 2026
The J-10C dogfight results represent a 'DeepSeek moment' for Chinese military tech, revealing far greater capability than most observers anticipated from a platform that had never before been proven in actual combat.
“Their deepseo or their deepseek moment where Chinese military tech which had never been proven in battle before demonstrated itself to be far more capable than a lot of people anticipated.”Michael Hirson · 15 May 2025
Despite Twitter chatter suggesting GLM is displacing DeepSeek, OpenCode usage data shows DeepSeek dips briefly at GLM launches but then bounces back.
“You can kind of see the DeepSeek one dip around the time GLM sort of comes out, but it's seems to sort of bounce back up after.”Jay V · 24 Jul 2026
Wall Street analysts are only now updating away from the DeepSeek-era assumption that Chinese labs sell frontier AI intelligence for pennies on the dollar.
“Wall Street analysts and people who aren't necessarily listening to the AI Daily Brief every day are finally beginning to update their priors and recognize that Chinese labs aren't, in most cases, selling frontier intelligence for pennies on the dollar the way they had previously assumed.”Nathaniel Whittemore · 18 Aug 2026