<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>The Minutes of Dylan Patel on gpu</title>
    <link>https://minutesof.com/dylan-patel/on/gpu/</link>
    <description>Everything Dylan Patel has said on gpu: 15 verbatim quotes between February 2023 and August 2026, each with a timestamp and a link to the recording it…</description>
    <language>en</language>
    <lastBuildDate>Sun, 30 Aug 2026 16:22:17 +0000</lastBuildDate>
    <atom:link href="https://minutesof.com/dylan-patel/on/gpu/feed.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Patel argues silicon supply can be optimized for either high throughput or high interactivity use cases.</title>
      <link>https://minutesof.com/q/6a3bd7b5-c74b-46a6-9dd5-8db72aa9576f/</link>
      <guid isPermaLink="true">https://minutesof.com/q/6a3bd7b5-c74b-46a6-9dd5-8db72aa9576f/</guid>
      <description>“Supply of silicon can go many ways. You can either leverage it to high throughput things or high interactivity things.” — SemiAnalysis</description>
      <pubDate>Mon, 17 Aug 2026 15:00:06 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel reveals SemiAnalysis operates over $80 million in compute for daily automated benchmarking across all ma</title>
      <link>https://minutesof.com/q/3047ee19-387d-4066-8826-f26d2fc2c5b3/</link>
      <guid isPermaLink="true">https://minutesof.com/q/3047ee19-387d-4066-8826-f26d2fc2c5b3/</guid>
      <description>“Every day it runs on an automated CI, and we run it on all the latest Chinese models from GLM, Zebu, Moonshot, Kimi, Alibaba, all these models we run.” — RAISE Summit</description>
      <pubDate>Thu, 16 Jul 2026 16:44:57 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>It&#x27;s 72 GPUs. It&#x27;s got more memory than Vera Rubin. It&#x27;s got more more, memory, flops. It&#x27;s got it&#x27;s it&#x27;s bett</title>
      <link>https://minutesof.com/q/ea805895-0928-4287-937a-0a6ccfad436b/</link>
      <guid isPermaLink="true">https://minutesof.com/q/ea805895-0928-4287-937a-0a6ccfad436b/</guid>
      <description>“It&#x27;s 72 GPUs. It&#x27;s got more memory than Vera Rubin. It&#x27;s got more more, memory, flops. It&#x27;s got it&#x27;s it&#x27;s better than Vera Rubin in most every way. It&#x27;s a little bit later.” — Supermicro</description>
      <pubDate>Wed, 15 Jul 2026 00:08:28 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel says AMD MI450X has 72 GPUs, more memory and flops than Vera Rubin, launching three to six months later.</title>
      <link>https://minutesof.com/q/250999e7-9869-45e6-95fa-d85ab1f347ac/</link>
      <guid isPermaLink="true">https://minutesof.com/q/250999e7-9869-45e6-95fa-d85ab1f347ac/</guid>
      <description>“It&#x27;s got more more, memory, flops. It&#x27;s got it&#x27;s it&#x27;s better than Vera Rubin in most every way. It&#x27;s a little bit later. I mean, three to six months after,” — Supermicro</description>
      <pubDate>Wed, 15 Jul 2026 00:08:28 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel reports Trainium rents for under $10 billion per gigawatt while GPUs cost $12-13 billion.</title>
      <link>https://minutesof.com/q/32312d2c-7043-4e82-9029-8dcbaeafc456/</link>
      <guid isPermaLink="true">https://minutesof.com/q/32312d2c-7043-4e82-9029-8dcbaeafc456/</guid>
      <description>“Tranium sells at sub $10,000,000,000 per gigawatt rental rate to Anthropic and to OpenAI. GPUs, at least before the craziness of the last six months, usually went around 12 to $13,000,000,000 per gigawatt.” — Sequoia Capital</description>
      <pubDate>Tue, 30 Jun 2026 12:00:25 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>We try and track the entire supply chain from tools that manufacture chips, fabs, data centers, energy, indust</title>
      <link>https://minutesof.com/q/8069a2f4-c1da-4a82-b880-1c456a111132/</link>
      <guid isPermaLink="true">https://minutesof.com/q/8069a2f4-c1da-4a82-b880-1c456a111132/</guid>
      <description>“We try and track the entire supply chain from tools that manufacture chips, fabs, data centers, energy, industrials, and then AI models and who&#x27;s using them, how much, and where.” — Swole as a Service</description>
      <pubDate>Wed, 13 May 2026 20:00:36 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel reports 10-15% of NVIDIA GPUs fail and require RMA within first two weeks of cluster deployment.</title>
      <link>https://minutesof.com/q/0dc25ccb-27dc-490b-bdd1-24ad4612e6ab/</link>
      <guid isPermaLink="true">https://minutesof.com/q/0dc25ccb-27dc-490b-bdd1-24ad4612e6ab/</guid>
      <description>“When you first turn on the cluster, about ten to fifteen percent of them fail RMA in the first two weeks. Wow. And then that&#x27;s fine. Like you have to receipt them, whatever.” — TBPN</description>
      <pubDate>Tue, 03 Feb 2026 23:08:03 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel says 10-15% of NVIDIA GPUs fail and need RMA in the first two weeks after cluster deployment.</title>
      <link>https://minutesof.com/q/267f5cee-22ee-4be6-bf05-893375b653c5/</link>
      <guid isPermaLink="true">https://minutesof.com/q/267f5cee-22ee-4be6-bf05-893375b653c5/</guid>
      <description>“When you first turn on the cluster, about ten to fifteen percent of them fail RMA in the first two weeks. Wow. And then that&#x27;s fine.” — TBPN</description>
      <pubDate>Tue, 03 Feb 2026 23:08:03 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel says OpenAI and Meta run NVIDIA GPUs at lower power to fit 10% more chips despite worse TCO.</title>
      <link>https://minutesof.com/q/d22ded74-0a82-4ccb-b2a1-0abee1378bda/</link>
      <guid isPermaLink="true">https://minutesof.com/q/d22ded74-0a82-4ccb-b2a1-0abee1378bda/</guid>
      <description>“Even though it&#x27;s terrible on a TCO basis, they were able to get, you know, 10% more GPUs in, and it&#x27;s great. And Meta has done similar.” — Open Compute Project</description>
      <pubDate>Thu, 23 Oct 2025 06:01:45 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel reveals DeepSeek inference implementation requires 160 GPUs worth over $10 million of hardware per repli</title>
      <link>https://minutesof.com/q/13cb1933-7756-463f-ac46-7f8cd066bee0/</link>
      <guid isPermaLink="true">https://minutesof.com/q/13cb1933-7756-463f-ac46-7f8cd066bee0/</guid>
      <description>“That&#x27;s over $10,000,000 of hardware, and then that&#x27;s just one replica, then you&#x27;ll have a lot of replicas and you share the caching servers between them.” — No Priors: AI, Machine Learning, Tech, &amp; Startups</description>
      <pubDate>Thu, 14 Aug 2025 10:01:35 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel says NVIDIA made 4 million GPUs last year and will produce 7 million this year.</title>
      <link>https://minutesof.com/q/08661a67-8385-4a47-bc96-ca62dcc63083/</link>
      <guid isPermaLink="true">https://minutesof.com/q/08661a67-8385-4a47-bc96-ca62dcc63083/</guid>
      <description>“Nvidia made over 4,000,000 GPUs last year, they&#x27;re making over 7,000,000 this year, right? High end data center GPUs.” — Special Competitive Studies Project</description>
      <pubDate>Thu, 13 Mar 2025 16:42:15 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel reports China imported one million H20 GPUs in recent quarters, enough for largest cluster.</title>
      <link>https://minutesof.com/q/2e9b3ee4-e698-4516-a8b0-920ae1a288c3/</link>
      <guid isPermaLink="true">https://minutesof.com/q/2e9b3ee4-e698-4516-a8b0-920ae1a288c3/</guid>
      <description>“Just in Q3, Q4 and the early part of Q1 this year, they imported a million H20s, Right?” — Special Competitive Studies Project</description>
      <pubDate>Thu, 13 Mar 2025 16:42:15 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel says GPU orders take four to six months from placement to data center installation.</title>
      <link>https://minutesof.com/q/fb13f24f-d903-4521-9be1-a668dcec5b78/</link>
      <guid isPermaLink="true">https://minutesof.com/q/fb13f24f-d903-4521-9be1-a668dcec5b78/</guid>
      <description>“call it four or five five, six months between, you know, when an order is placed and you can actually have it installed in your data center if it got worked on immediately.” — The Inside View</description>
      <pubDate>Wed, 09 Aug 2023 15:00:23 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel says NVIDIA GPU bandwidth increased less than 10x while FLOPS increased 100x from 2016 to 2023.</title>
      <link>https://minutesof.com/q/45f2aadd-215b-4a37-97e7-67d1873d78a9/</link>
      <guid isPermaLink="true">https://minutesof.com/q/45f2aadd-215b-4a37-97e7-67d1873d78a9/</guid>
      <description>“The bandwidth has not even gone up one order of magnitude. Right? Whereas flops have gone up two orders of magnitude. So so less than one order of magnitude versus two orders of magnitude increase.” — Gradient Flow</description>
      <pubDate>Thu, 02 Feb 2023 12:00:39 +0000</pubDate>
      <category>gpu</category>
    </item>
    <item>
      <title>Patel states memory cost has quadrupled in NVIDIA GPUs while per-gigabyte pricing remained flat since 2016.</title>
      <link>https://minutesof.com/q/fad9d07a-a3d6-43a9-9e0b-f2591b6d3829/</link>
      <guid isPermaLink="true">https://minutesof.com/q/fad9d07a-a3d6-43a9-9e0b-f2591b6d3829/</guid>
      <description>“And now with the h 100, they have 96 or 80 gigabytes of memory, but the cost per gigabyte is the same.” — Gradient Flow</description>
      <pubDate>Thu, 02 Feb 2023 12:00:39 +0000</pubDate>
      <category>gpu</category>
    </item>
  </channel>
</rss>
