On the record about
2 people · 5 quotes · 13 Oct 2024 to 30 Jun 2026
2 of 2 lanes rest on fewer than 5 quotes and are marked thin. Offsets are days from the middle first-quote date, 13 Oct 2024 — a date, and nothing else. It is not a claim about who reached a view first.
Gurley notes early internet startups shifted from 100% Oracle and Sun to Linux and MySQL in five years.
“Every single fucking one of them were running on Oracle and Sun. And five years later, they were all running on Linux and MySQL, like in five years.”
Patel states avoiding redundant prefill through KV cache can cut inference costs to one-fourth of previous levels.
“So you can cut your cost to one fourth of what it was previously if you just don't do the pre fill. Right?”
Patel describes simultaneous gains across pre-training, scaling, quantization, systems, and RL infrastructure and methods.
“There's tons of gains in quantization and systems, but there's also tons of gains in RL stuff and both the, like, you know, infra side of things and non infra side of things.”
Gurley predicts AI will shift from innovation to optimization like dot-com era Oracle-to-Linux migration.
“So there were you went from a period of innovation to a period of optimization where people are much more worried about cost.”
Patel argues co-optimizing hardware, software, and models produces 100x gains instead of 8x from multiplicative 2x improvements per layer.
“The real breakthrough innovation is when you leapfrog a few layers, you co optimize and co design them, and now all of a sudden you've taken what could have been a 2x here, 2x here, 2x here, and instead of being multiplicative to 8x, it's actually 100x, because you've optimized it across all three layers.”