Brad Gerstner on

ai inference

2 entries, 28 Aug 2025 to 13 Jul 2026

On the recordsourced and dated, oldest first

    1. spoken

      Gerstner reports Google inference generation increased 100x in a year, from 9 trillion to 980 trillion tokens monthly.

      “A year ago, Google per month was doing about 9,000,000,000,000 tokens a month in terms of inference generation, right, compute generation. Today, it's 980,000,000,000,000 tokens.”

      28 Aug 2025 · CNBC Television · 2:19 · source · permalink
    1. spoken

      Gerstner argues the difference between $3 and $15 inference costs is irrelevant when replacing a $200 per hour consultant.

      “The difference between spending $3 on a cheap model or $15 on an expensive model to replace a $200 an hour consultant. It's just irrelevance. That inference cost difference is irrelevant”

      13 Jul 2026 · All-In Podcast · 0:50 · source · permalink

Brad Gerstner ontheir other subjects

Everything Brad Gerstner is on record saying · RSS