<?xml version="1.0" encoding="UTF-8"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
  <channel>
    <title>The Minutes of Bill Gurley on inference</title>
    <link>https://minutesof.com/bill-gurley/on/inference/</link>
    <description>Everything Bill Gurley has said on inference: 3 verbatim quotes between September 2024 and December 2024, each with a timestamp and a link to the…</description>
    <language>en</language>
    <lastBuildDate>Sun, 30 Aug 2026 16:22:02 +0000</lastBuildDate>
    <atom:link href="https://minutesof.com/bill-gurley/on/inference/feed.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Patel describes reasoning models generating thousands of thinking tokens, sometimes switching languages mid-re</title>
      <link>https://minutesof.com/q/476ea39f-bfb6-495d-9a59-81ea69699b0d/</link>
      <guid isPermaLink="true">https://minutesof.com/q/476ea39f-bfb6-495d-9a59-81ea69699b0d/</guid>
      <description>“It generates tons of things. It&#x27;s like it it sometimes switches between Chinese and English. Right? Like, whatever it is. It&#x27;s thinking. Right? It&#x27;s churning.” — Bg2 Pod</description>
      <pubDate>Mon, 23 Dec 2024 20:25:20 +0000</pubDate>
      <category>inference</category>
    </item>
    <item>
      <title>Patel explains reasoning models increase cost ten times by outputting 11,000 tokens versus 1,000 for same quer</title>
      <link>https://minutesof.com/q/9aed22e0-c898-4e8e-9a74-966679c754b1/</link>
      <guid isPermaLink="true">https://minutesof.com/q/9aed22e0-c898-4e8e-9a74-966679c754b1/</guid>
      <description>“I outputted a thousand tokens to I outputted 11,000 tokens. I&#x27;ve 10x&#x27;d my spend to generate no. Not the same thing. Right? It&#x27;s higher quality.” — Bg2 Pod</description>
      <pubDate>Mon, 23 Dec 2024 20:25:20 +0000</pubDate>
      <category>inference</category>
    </item>
    <item>
      <title>Patel calculates reasoning models cost fifty times more per query due to batch size and token generation combi</title>
      <link>https://minutesof.com/q/3af95131-05a1-45d4-ae2d-9578bd46190d/</link>
      <guid isPermaLink="true">https://minutesof.com/q/3af95131-05a1-45d4-ae2d-9578bd46190d/</guid>
      <description>“Cost increase for a single token to be generated is four to five x, but then I&#x27;m generating 10 x as many tokens.” — Bg2 Pod</description>
      <pubDate>Mon, 23 Dec 2024 20:25:20 +0000</pubDate>
      <category>inference</category>
    </item>
  </channel>
</rss>
