<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>deepseek: everyone on the record — The Minutes</title>
    <link>https://minutesof.com/on/deepseek/</link>
    <description>Everyone on the record about deepseek: 21 verbatim quotes from 5 people, January 2025 to June 2026, in one chronology, each with a timestamp and a link to…</description>
    <language>en</language>
    <lastBuildDate>Sun, 30 Aug 2026 22:32:55 +0000</lastBuildDate>
    <atom:link href="https://minutesof.com/on/deepseek/feed.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Dylan Patel: Patel claims hardware improved 30x from Hopper to Blackwell for DeepSeek on optimized deployments</title>
      <link>https://minutesof.com/q/1aeaaf0c-1080-4d21-a79c-2bc3d127baf5/</link>
      <guid isPermaLink="true">https://minutesof.com/q/1aeaaf0c-1080-4d21-a79c-2bc3d127baf5/</guid>
      <description>“from Hopper to Blackwell, which is all we&#x27;ve had over the last three years, roughly 30x improvement on DeepSeek, on the most optimized deployment, which you can see on InferenceX there&#x27;s about a 30x improvement.” — Dylan Patel, Sequoia Capital</description>
      <pubDate>Tue, 30 Jun 2026 12:00:25 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Brad Gerstner: Gerstner says Chinese AI stack including DeepSeek and other models is making disturbing inroads</title>
      <link>https://minutesof.com/q/0747bd60-bb55-46d1-98e7-3331e41198ed/</link>
      <guid isPermaLink="true">https://minutesof.com/q/0747bd60-bb55-46d1-98e7-3331e41198ed/</guid>
      <description>“No. For sure. And in fact, I I hear disturbing things all the time about the Chinese stack making great inroads in The Middle East and in the Southern Hemisphere, etcetera, because, frankly, they&#x27;re working really quickly to put their chips and their models, their open source models, Deepsea, Kimi k two, Quinn, etcetera, and exporting those to the world.” — Brad Gerstner, Altimeter Capital</description>
      <pubDate>Tue, 10 Feb 2026 21:40:23 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel cites DeepSeek&#x27;s open-sourced inference system requiring 140 GPUs communicating over RDMA n</title>
      <link>https://minutesof.com/q/b5b15eca-a046-4a24-8309-0cfa9d92ceec/</link>
      <guid isPermaLink="true">https://minutesof.com/q/b5b15eca-a046-4a24-8309-0cfa9d92ceec/</guid>
      <description>“One example is one that DeepSeek open sourced over December of last yearJanuary, February of this year, where a single replica of inference for a single model is going to be like 140 GPUs.” — Dylan Patel, Clockwork</description>
      <pubDate>Fri, 21 Nov 2025 17:18:51 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel reveals DeepSeek inference implementation requires 160 GPUs worth over $10 million of hardw</title>
      <link>https://minutesof.com/q/13cb1933-7756-463f-ac46-7f8cd066bee0/</link>
      <guid isPermaLink="true">https://minutesof.com/q/13cb1933-7756-463f-ac46-7f8cd066bee0/</guid>
      <description>“That&#x27;s over $10,000,000 of hardware, and then that&#x27;s just one replica, then you&#x27;ll have a lot of replicas and you share the caching servers between them.” — Dylan Patel, No Priors: AI, Machine Learning, Tech, &amp; Startups</description>
      <pubDate>Thu, 14 Aug 2025 10:01:35 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel says DeepSeek&#x27;s cost efficiency follows the expected trend line, just from an unexpected so</title>
      <link>https://minutesof.com/q/42277ecd-5d1b-4270-a90f-11eb7f3a67b8/</link>
      <guid isPermaLink="true">https://minutesof.com/q/42277ecd-5d1b-4270-a90f-11eb7f3a67b8/</guid>
      <description>“It&#x27;s not unexpected. Right? Like, this is actually within the trend line of what happened with GPT three is happening to GPT four level quality with DeepSeq.” — Dylan Patel, Alex Kantrowitz</description>
      <pubDate>Wed, 23 Apr 2025 16:30:06 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel argues the surprise was a Chinese company achieving this cost reduction, not the reduction</title>
      <link>https://minutesof.com/q/a64a56f5-df99-4993-bd99-c6e00a54110d/</link>
      <guid isPermaLink="true">https://minutesof.com/q/a64a56f5-df99-4993-bd99-c6e00a54110d/</guid>
      <description>“I think what was really surprising was that it was a Chinese company for the first time. Right? Because Google and and OpenAI and Anthropic and Meta have all traded blows.” — Dylan Patel, Alex Kantrowitz</description>
      <pubDate>Wed, 23 Apr 2025 16:30:06 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel reports GPT-3 to Llama 3.2 costs fell 1,200x while GPT-4 to DeepSeek v3 costs fell 600x.</title>
      <link>https://minutesof.com/q/c28a19fc-d8f5-4e4a-bc59-e7dabb126c8c/</link>
      <guid isPermaLink="true">https://minutesof.com/q/c28a19fc-d8f5-4e4a-bc59-e7dabb126c8c/</guid>
      <description>“when we looked at g p d three, the cost fell 1,200 x from g p d three&#x27;s initial cost to what you can get Lama 3.23 b today.” — Dylan Patel, Alex Kantrowitz</description>
      <pubDate>Wed, 23 Apr 2025 16:30:06 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel challenges DeepSeek&#x27;s GPU count claims, citing job ads promising tens of thousands of GPUs.</title>
      <link>https://minutesof.com/q/c235c852-3da5-495d-8118-6bceb59270eb/</link>
      <guid isPermaLink="true">https://minutesof.com/q/c235c852-3da5-495d-8118-6bceb59270eb/</guid>
      <description>“The estimates that we have is, so first of all, ads in China, they say they have tens of thousands of GPUs for researchers, right?” — Dylan Patel, Special Competitive Studies Project</description>
      <pubDate>Thu, 13 Mar 2025 16:42:15 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Bill Gurley: Gurley reports 1,300 variants of DeepSeek R1 appeared on Hugging Face within days of release.</title>
      <link>https://minutesof.com/q/2fad991a-2404-440c-9dd6-13bf493b7e2b/</link>
      <guid isPermaLink="true">https://minutesof.com/q/2fad991a-2404-440c-9dd6-13bf493b7e2b/</guid>
      <description>“forty eight hours after r one was posted, they had 500 variants on Hugging Face. And today I pinged them this morning before we started, they&#x27;re up to 1,300.” — Bill Gurley, BG2 Pod</description>
      <pubDate>Wed, 05 Feb 2025 23:13:28 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Brad Gerstner: Gerstner reports DeepSeek is massively compute constrained now, with GPU differential growing v</title>
      <link>https://minutesof.com/q/7915177a-c38e-49f6-8c41-f1704a1047cb/</link>
      <guid isPermaLink="true">https://minutesof.com/q/7915177a-c38e-49f6-8c41-f1704a1047cb/</guid>
      <description>“On top of that, we learned, you know, they are massively compute constrained right now. So you&#x27;ve seen some tweets about this, people that are, you know, getting server delay and all this stuff” — Brad Gerstner, BG2 Pod</description>
      <pubDate>Wed, 05 Feb 2025 23:13:28 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Bill Gurley: Gurley reports DeepSeek pricing is one twentieth of OpenAI on comparable API models.</title>
      <link>https://minutesof.com/q/f2fcae4d-e3e9-474b-9572-95f3a2e55fad/</link>
      <guid isPermaLink="true">https://minutesof.com/q/f2fcae4d-e3e9-474b-9572-95f3a2e55fad/</guid>
      <description>“if you look at the models that are apples to apples on the API right now, that that DeepSeek&#x27;s pricing about one twentieth of OpenAI.” — Bill Gurley, BG2 Pod</description>
      <pubDate>Wed, 05 Feb 2025 23:13:28 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Bill Gurley: Gurley explains DeepSeek innovated by separating parameters to work with smaller counts faster th</title>
      <link>https://minutesof.com/q/dd419e62-f500-4370-8d57-c84a06675948/</link>
      <guid isPermaLink="true">https://minutesof.com/q/dd419e62-f500-4370-8d57-c84a06675948/</guid>
      <description>“They were able to do that because they figured out a way to separate the parameters and work on things with smaller parameter counts faster, which no one else had done before.” — Bill Gurley, BG2 Pod</description>
      <pubDate>Wed, 05 Feb 2025 23:13:28 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel notes DeepSeek&#x27;s routing innovation removing auxiliary loss represents compounding small im</title>
      <link>https://minutesof.com/q/fab5c10d-e2cd-4e52-b056-9586751d1efb/</link>
      <guid isPermaLink="true">https://minutesof.com/q/fab5c10d-e2cd-4e52-b056-9586751d1efb/</guid>
      <description>“this type of change can be big, it can be small, but they add up over time.” — Dylan Patel, Lex Fridman</description>
      <pubDate>Mon, 03 Feb 2025 00:12:13 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel says DeepSeek modifies code at or below NVIDIA&#x27;s CUDA layer, a rare technical capability.</title>
      <link>https://minutesof.com/q/99d651cc-75d0-4621-9e78-21bd90eeb8aa/</link>
      <guid isPermaLink="true">https://minutesof.com/q/99d651cc-75d0-4621-9e78-21bd90eeb8aa/</guid>
      <description>“For example, on their to get highly efficient training, they&#x27;re making modifications at or below the CUDA layer for NVIDIA chips.” — Dylan Patel, Lex Fridman</description>
      <pubDate>Mon, 03 Feb 2025 00:12:13 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Dylan Patel: Patel explains DeepSeek v3 base is trained once, then post-trained differently to create chat ver</title>
      <link>https://minutesof.com/q/02fd4a40-e301-43f7-be83-33801b8010a9/</link>
      <guid isPermaLink="true">https://minutesof.com/q/02fd4a40-e301-43f7-be83-33801b8010a9/</guid>
      <description>“This reasoning model has a lot of overlapping training steps to DeepSeek v three, and it&#x27;s confusing that you have a base model called v three that you do something to to get a chat model, and then you do some different things to get a reasoning model.” — Dylan Patel, Lex Fridman</description>
      <pubDate>Mon, 03 Feb 2025 00:12:13 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>David Sacks: Sacks says DeepSeek v3 self-identified as ChatGPT-4 five out of eight times when asked.</title>
      <link>https://minutesof.com/q/e157002e-0548-4018-aec6-14d72fa3e027/</link>
      <guid isPermaLink="true">https://minutesof.com/q/e157002e-0548-4018-aec6-14d72fa3e027/</guid>
      <description>“When you would ask it, who are you? Like what model are you? Five out of eight times, v three would tell you that it was ChatGPT four.” — David Sacks, All-In Podcast</description>
      <pubDate>Fri, 31 Jan 2025 21:52:00 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>David Sacks: Sacks says DeepSeek&#x27;s compute cluster costs over $1 billion, contradicting the $6 million narrati</title>
      <link>https://minutesof.com/q/c35e5045-0094-4d11-8d15-94e56f582513/</link>
      <guid isPermaLink="true">https://minutesof.com/q/c35e5045-0094-4d11-8d15-94e56f582513/</guid>
      <description>“you add up the the cost of a compute cluster with 50,000 plus hoppers and it&#x27;s gonna be over $1,000,000,000.” — David Sacks, All-In Podcast</description>
      <pubDate>Fri, 31 Jan 2025 21:52:00 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>David Sacks: Sacks cites Dylan Patel&#x27;s estimate that DeepSeek has 50,000 Hopper chips across different models.</title>
      <link>https://minutesof.com/q/a4ede3f9-51af-4927-b48a-d925bdf7fed7/</link>
      <guid isPermaLink="true">https://minutesof.com/q/a4ede3f9-51af-4927-b48a-d925bdf7fed7/</guid>
      <description>“Dylan Patel, who&#x27;s leading semiconductor analyst, has estimated that DeepSeac has about 50,000 hoppers. And specifically, he said they have about 10,000 h one hundreds, They have 10,000 h eight hundreds and 30,000 h twenties.” — David Sacks, All-In Podcast</description>
      <pubDate>Fri, 31 Jan 2025 21:52:00 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>David Sacks: Sacks says the $6 million DeepSeek training cost claim should be debunked, corroborated by indust</title>
      <link>https://minutesof.com/q/73cecf18-6302-4ffe-ba6b-32c73ed695f1/</link>
      <guid isPermaLink="true">https://minutesof.com/q/73cecf18-6302-4ffe-ba6b-32c73ed695f1/</guid>
      <description>“On this one, I&#x27;m with Palmer Luckey and Brad Gerstner and others, and I think this has been pretty much corroborated by everyone I&#x27;ve talked to that that number should be debunked.” — David Sacks, All-In Podcast</description>
      <pubDate>Fri, 31 Jan 2025 21:52:00 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>David Sacks: Sacks says people would have been surprised a Chinese company released the second reasoning model</title>
      <link>https://minutesof.com/q/62b103f6-6774-4c42-9053-300c3f5ea97c/</link>
      <guid isPermaLink="true">https://minutesof.com/q/62b103f6-6774-4c42-9053-300c3f5ea97c/</guid>
      <description>“if you had said to people a few weeks ago that the second company to release a reasoning model along the lines of o one would be a Chinese company, I think people would have been surprised by that.” — David Sacks, All-In Podcast</description>
      <pubDate>Fri, 31 Jan 2025 21:52:00 +0000</pubDate>
      <category>deepseek</category>
    </item>
    <item>
      <title>Gavin Baker: Baker predicts frontier AI labs will stop releasing leading-edge models to prevent IP theft via k</title>
      <link>https://minutesof.com/q/5fe5835f-3bc6-4b11-b3f5-6a03e58df6ec/</link>
      <guid isPermaLink="true">https://minutesof.com/q/5fe5835f-3bc6-4b11-b3f5-6a03e58df6ec/</guid>
      <description>“I think you will see the Frontier Labs stop releasing their leading edge models to prevent knowledge distillation and their IP effectively being stolen.” — Gavin Baker, All-In Podcast</description>
      <pubDate>Sat, 04 Jan 2025 00:23:00 +0000</pubDate>
      <category>deepseek</category>
    </item>
  </channel>
</rss>
