Dylan Patel on

cost

4 entries, 23 Dec 2024 to 22 Jul 2026

On the recordsourced and dated, oldest first

    1. spoken

      Patel calculates o1 models increase inference costs by roughly 50x compared to standard models per query.

      “Cost increase for a single token to be generated is four to five x, but then I'm generating 10 x as many tokens.”

      23 Dec 2024 · BG2 Pod · 50:06 · source · permalink
    1. spoken

      Patel says DeepSeek's model is 400 to 600x cheaper than original GPT-4.

      “And so now DeepSeek's model is something like 400 to 600x cheaper than the original GPT-four.”

      27 Mar 2025 · MedBricks Webcast · 31:36 · source · permalink
    2. spoken

      Patel states each gigawatt AI deployment costs tens of billions of dollars end-to-end.

      “Each of these gigawatt deployments once you take it from all the way from the power to the GPUs and the cloud contract, those are tens of billions of dollars each gigawatt.”

      2 Sep 2025 · Nebius · 28:43 · source · permalink
    1. spoken

      Patel reports that tokens now represent 30% of his 90-person company's costs versus 70% for employees.

      “My my own company of 90 people, 30% of my cost now is tokens versus 70% employee costs.”

      22 Jul 2026 · RAISE Summit · 7:23 · source · permalink

Dylan Patel ontheir other subjects

Everything Dylan Patel is on record saying · RSS