<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>The Minutes of Dylan Patel on scaling</title>
    <link>https://minutesof.com/dylan-patel/on/scaling/</link>
    <description>Everything Dylan Patel has said on scaling: 5 verbatim quotes between December 2023 and August 2026, each with a timestamp and a link to the recording it…</description>
    <language>en</language>
    <lastBuildDate>Sun, 30 Aug 2026 16:22:17 +0000</lastBuildDate>
    <atom:link href="https://minutesof.com/dylan-patel/on/scaling/feed.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Patel reports Microsoft aims to 10x training capacity every 18-24 months, representing a 10x increase from GPT</title>
      <link>https://minutesof.com/q/fa0ed250-859f-4800-a571-4fb9a5440edb/</link>
      <guid isPermaLink="true">https://minutesof.com/q/fa0ed250-859f-4800-a571-4fb9a5440edb/</guid>
      <description>“We try to 10x the training capacity every eighteen to twenty four months. And so this would be effectively a 10x increase. 10x from what GPD five was trained with.” — Dwarkesh Patel</description>
      <pubDate>Wed, 12 Nov 2025 17:02:10 +0000</pubDate>
      <category>scaling</category>
    </item>
    <item>
      <title>Patel states $10 billion data centers target automated software engineering, not chat models.</title>
      <link>https://minutesof.com/q/119203f3-2761-4bf3-a619-aaad40c2f71a/</link>
      <guid isPermaLink="true">https://minutesof.com/q/119203f3-2761-4bf3-a619-aaad40c2f71a/</guid>
      <description>“So no one is trying to make with these $10,000,000,000 data centers, they&#x27;re not trying to make chat models. Right?” — Alex Kantrowitz</description>
      <pubDate>Wed, 23 Apr 2025 16:30:06 +0000</pubDate>
      <category>scaling</category>
    </item>
    <item>
      <title>Patel says current models use 100,000 GPUs while next generation will require hundreds of thousands or million</title>
      <link>https://minutesof.com/q/561bd369-2c88-4ea1-8fa4-3d02389be50d/</link>
      <guid isPermaLink="true">https://minutesof.com/q/561bd369-2c88-4ea1-8fa4-3d02389be50d/</guid>
      <description>“And next generation models that are trained on hundreds of thousands or even millions GPUs, right?” — Special Competitive Studies Project</description>
      <pubDate>Thu, 13 Mar 2025 16:42:15 +0000</pubDate>
      <category>scaling</category>
    </item>
    <item>
      <title>Patel calculates next-generation clusters deliver 15x more compute through five times more GPUs and 3x perform</title>
      <link>https://minutesof.com/q/badb57b8-ea1b-45cb-887a-c8b1bba9f5d8/</link>
      <guid isPermaLink="true">https://minutesof.com/q/badb57b8-ea1b-45cb-887a-c8b1bba9f5d8/</guid>
      <description>“So you got you have five x the GPUs, and you have three x the performance per GPU roughly. So then you&#x27;re at, like, 15 x more compute.” — Unsupervised Learning: With Jacob Effron</description>
      <pubDate>Tue, 21 Jan 2025 14:00:12 +0000</pubDate>
      <category>scaling</category>
    </item>
    <item>
      <title>Patel calculates that multi-trillion parameter models require transmitting 40 terabytes of data every two seco</title>
      <link>https://minutesof.com/q/2b1dd77d-4b29-490c-a1f9-fd32caff427b/</link>
      <guid isPermaLink="true">https://minutesof.com/q/2b1dd77d-4b29-490c-a1f9-fd32caff427b/</guid>
      <description>“They&#x27;re doing it for like multi trillion. Right? So let&#x27;s call it 10,000,000,000,000 parameters, four bytes parameter, that&#x27;s 40 terabytes of data you need to transmit in two seconds.” — Scaling Intelligence</description>
      <pubDate>Tue, 12 Nov 2024 04:17:00 +0000</pubDate>
      <category>scaling</category>
    </item>
  </channel>
</rss>
