<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>The Minutes of Dylan Patel on gb200</title>
    <link>https://minutesof.com/dylan-patel/on/gb200/</link>
    <description>Everything Dylan Patel has said on gb200: 5 verbatim quotes from October 2025, each with a timestamp and a link to the recording it came from.</description>
    <language>en</language>
    <lastBuildDate>Sun, 30 Aug 2026 16:22:19 +0000</lastBuildDate>
    <atom:link href="https://minutesof.com/dylan-patel/on/gb200/feed.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Patel quantifies GB200 as 10x more power efficient than H200 at certain interactivity rates.</title>
      <link>https://minutesof.com/q/5246be24-b713-4f40-9330-362372209a1f/</link>
      <guid isPermaLink="true">https://minutesof.com/q/5246be24-b713-4f40-9330-362372209a1f/</guid>
      <description>“at certain interactivity rates, I. E. Tokens per second per user, it&#x27;s 10x more efficient per watt, right, compared to H200.” — Open Compute Project</description>
      <pubDate>Thu, 23 Oct 2025 06:01:19 +0000</pubDate>
      <category>gb200</category>
    </item>
    <item>
      <title>Patel notes GB200 expanded NVLink from 8 to 72 GPUs and rack power jumped from 10 to 140 kilowatts.</title>
      <link>https://minutesof.com/q/015334ac-c584-41cb-9262-58aa5ef55a91/</link>
      <guid isPermaLink="true">https://minutesof.com/q/015334ac-c584-41cb-9262-58aa5ef55a91/</guid>
      <description>“Now you have 72 GPUs. And if you go look at the rack, right? It&#x27;s completely different, right? It&#x27;s liquid cooled. It&#x27;s an entire rack that consumes 140 kilowatts, whereas H100 servers consumed 10 kilowatts.” — Together AI</description>
      <pubDate>Fri, 03 Oct 2025 20:23:30 +0000</pubDate>
      <category>gb200</category>
    </item>
    <item>
      <title>Patel explains GB200&#x27;s 72-GPU configuration creates reliability challenges compared to 8-GPU systems due to hi</title>
      <link>https://minutesof.com/q/e442c5ed-ed97-4cb7-8195-b9012b3b0fc8/</link>
      <guid isPermaLink="true">https://minutesof.com/q/e442c5ed-ed97-4cb7-8195-b9012b3b0fc8/</guid>
      <description>“If the reliability of each GPU is the same, then when a single GPU fails in 72 GPUs, you have a much higher chance of something failing, right?” — Together AI</description>
      <pubDate>Fri, 03 Oct 2025 20:23:30 +0000</pubDate>
      <category>gb200</category>
    </item>
    <item>
      <title>Patel confirms OpenAI runs production inference on GB200 despite reliability requiring workloads handle 64 of</title>
      <link>https://minutesof.com/q/63137c12-d683-47eb-9bf7-16977addf85c/</link>
      <guid isPermaLink="true">https://minutesof.com/q/63137c12-d683-47eb-9bf7-16977addf85c/</guid>
      <description>“OpenAI has said they&#x27;re running production inference on GV200 a couple of months ago, in fact. Right?” — Together AI</description>
      <pubDate>Fri, 03 Oct 2025 20:23:30 +0000</pubDate>
      <category>gb200</category>
    </item>
    <item>
      <title>Patel says B200 is better for training while GB200 is better for inference, reversing expected use.</title>
      <link>https://minutesof.com/q/b34c2dfe-463a-4923-8ba7-41d8b89b38a6/</link>
      <guid isPermaLink="true">https://minutesof.com/q/b34c2dfe-463a-4923-8ba7-41d8b89b38a6/</guid>
      <description>“And so you&#x27;ve sort of Which is the exact opposite of what you would have expected. Oh, use the big thing for training and use the small thing for inference.” — Together AI</description>
      <pubDate>Fri, 03 Oct 2025 20:23:30 +0000</pubDate>
      <category>gb200</category>
    </item>
  </channel>
</rss>
