On the record about

optimization

2 people · 5 quotes · 13 Oct 2024 to 30 Jun 2026

Who is on this subjectordered by the date of their first quote here

2 of 2 lanes rest on fewer than 5 quotes and are marked thin. Offsets are days from the middle first-quote date, 13 Oct 2024 — a date, and nothing else. It is not a claim about who reached a view first.

The chronologysourced and dated, oldest first

    1. Bill Gurley

      Gurley notes early internet startups shifted from 100% Oracle and Sun to Linux and MySQL in five years.

      “Every single fucking one of them were running on Oracle and Sun. And five years later, they were all running on Linux and MySQL, like in five years.”

      13 Oct 2024 · BG2 Pod · 24:32 · source · permalink
    1. Dylan Patel

      Patel states avoiding redundant prefill through KV cache can cut inference costs to one-fourth of previous levels.

      “So you can cut your cost to one fourth of what it was previously if you just don't do the pre fill. Right?”

      21 Nov 2025 · Clockwork · 54:03 · source · permalink
    1. Dylan Patel

      Patel describes simultaneous gains across pre-training, scaling, quantization, systems, and RL infrastructure and methods.

      “There's tons of gains in quantization and systems, but there's also tons of gains in RL stuff and both the, like, you know, infra side of things and non infra side of things.”

      15 Jan 2026 · SAIL Media · 2:51 · source · permalink
    2. Bill Gurley

      Gurley predicts AI will shift from innovation to optimization like dot-com era Oracle-to-Linux migration.

      “So there were you went from a period of innovation to a period of optimization where people are much more worried about cost.”

      26 Jan 2026 · Yahoo Finance · 10:34 · source · permalink
    3. Dylan Patel

      Patel argues co-optimizing hardware, software, and models produces 100x gains instead of 8x from multiplicative 2x improvements per layer.

      “The real breakthrough innovation is when you leapfrog a few layers, you co optimize and co design them, and now all of a sudden you've taken what could have been a 2x here, 2x here, 2x here, and instead of being multiplicative to 8x, it's actually 100x, because you've optimized it across all three layers.”

      30 Jun 2026 · Sequoia Capital · 28:20 · source · permalink

Every subject on the record · RSS