Baker argues contracted compute trades at massive discount to spot, repricing will accelerate cash flows and answer ROI questions.
“And so, essentially, you have the contracted base of installed compute trading at a massive discount to the current spot market.”
Baker reports GPU rental prices doubled from mid-$2 to nearly $4 per hour over seven months for identical clusters.
“And they had rented a cluster of several thousand black wells, and we'll just call it somewhere in the mid $2 per GPU hour.”
Baker reports inference cloud expects to pay 100% more for Blackwells when contracts expire, showing hyperscalers are underearning.
“They went on a podcast, and they essentially said, we are planning to pay 100% more for Blackwell's when our contract expires. And that just means that essentially all the hyperscalers are under earning.”
Baker cites analysis showing compute margins, quantity, and inference margins all rising simultaneously, driving lab acceleration.
“The amount of compute is going up and inference margins going up. And if you multiply those three, that's how you're getting this crazy acceleration into some of the labs plus open source,”
Baker estimates SpaceX monetizes compute at $50B per gigawatt versus $73B consensus, with Grok and Cursor hitting $10B ARR quickly.
“And they're monetizing at something like 50,000,000,000 a gig and consensus estimates for next year are 73,000,000,000. So forget Starlink v three, forget Starlink direct to sell, Grok 4.5 and Cursor.”
Baker believes Anthropic is likely already generating cash or will start this year.
“I think Anthropic probably starts generating cash this year if they are not already generating cash, which I think is probably the case.”
Baker argues America will consume all available compute, reducing edge AI bear case concerns.
“And I just think the same is true of compute. It's why I'm probably less worried about like an edge AI bear case than I was.”
Baker explains harness engineering matters significantly and harnesses are increasingly co-developed with models.
“And it turns out that harness engineering is not as important as the model, but it really matters. And these harnesses in these models are increasingly being co developed.”
Baker says understanding frontier AI now requires enterprise usage-based plans, not consumer subscriptions, due to rate limiting.
“To understand what Frontier AI is capable of today, even for a non coding use case, need to have Cloud Code or Codex five point Codex. And you need to be on an enterprise plan.”
Baker says AI shifting from flat pricing to usage-based is extremely bullish as people consume more AI.
“AI is just shifting from all you can eat to pay by the drink. Then it turns out people really like to talk to their friends long distance.”
Baker predicts OpenAI and Anthropic will exceed $200B ARR this year due to shift to usage-based pricing.
“So I think the shift to usage based pricing is probably why you will see OpenAI and Anthropic exceed well over $200,000,000,000 in ARR this year.”
Baker predicts OpenAI and Anthropic will exceed $200B ARR this year due to usage-based pricing shift.
“I think the shift to usage based pricing is probably why you will see OpenAI and Anthropic exceed well over $200,000,000,000 in ARR this year.”
Baker argues GPU useful lives will extend to 10-15 years due to inference disaggregation, contradicting AI skeptics.
“The disaggregation of inference means that I think these GPUs are going to have ten or fifteen year lives. The AI skeptics are like, oh, these companies are all cooking their books.”
Baker predicts GPUs will have 10-15 year useful lives due to prefill-inference disaggregation, extending older chips' value.
“The disaggregation of inference means that I think these GPUs are going to have ten or fifteen year lives.”
Baker argues GPU useful lives will extend to 10-15 years due to prefill/decode disaggregation, not 1-2 years.
“The useful life of GPU is only a year or two. The useful life of CPU is only four years because the rapid technological change.”
Baker says AI models shifting to usage-based pricing with overage reveals no ceiling on spending yet.
“We're just moving from these all you can eat plans to usage based plans with overage, where those usage tokens cost a lot more, and we're finding out that there's we're nowhere near the amount of, you know, people ceiling price for how much they'll spend.”
Baker emphasizes only 0.1% of the world uses AI models properly yet there's massive shortage despite trillions spent.
“And we're in an insane shortage despite spending cumulatively trillions of dollars. What happens when 5% of the world's population is using these models the way the cutting edge 10 basis points are?”
Sohn Conference Foundation
“What happens when 5% of the world's population is using these models the way the cutting edge 10 basis points are? Like, it's just it's unimaginable. This is why orbital compute is a necessity.”
Sohn Conference Foundation
“Tranium, by far. Tranium is going to be to 2026, especially in the second half of this year when Tranium three really ramps, as TPUs were to twenty twenty five.”
Baker predicts Trainium will dominate 2026 like TPUs did in 2025, with Trainium 3 ramping in second half.
“Tranium is going to be to 2026, especially in the second half of this year when Tranium three really ramps, as TPUs were to twenty twenty five.”
Baker argues Trainium is most underestimated because frontier mixture-of-expert models require switched scale-up networks for inference.
“And so Trainium is for sure the most underestimated, not only because of those design choices, but because the all of these frontier models are what are called mixture of expert models.”
Baker states only NVIDIA and Amazon Trainium have functioning switched scale-up networks for inference today.
“And the only two functioning switched scale up networks in the world today are the ones that power NVIDIA GPUs and Amazon's Trainiums.”
Baker reveals Atreides could have invested over $50 million in CoreWeave at $1.1 billion valuation but was conflicted out.
“I could've Atreides could've invested over $50,000,000 in the round at 1,100,000,000, And I was conflicted out by Crusoe,”
Baker identifies three axes of AI scaling: pretraining, inference time compute, and now reasoning as the third multiplicative dimension.
“And then we started scaling around inference time compute. And it's very clear that we have now added a third axis of scaling performance, and that is reasoning.”
Baker predicts Tesla FSD will achieve 100x improvement quickly as compute scales to GPT-4.5 levels.
“I think they're going to go really fast to GPT-4.5 compute, which means you're going to get, using these orders of magnitude, you're going get a 100x improvement really fast.”
Baker argues AI revolution stems from cloud computing power and mobile-generated data, not algorithmic advances.
“The only thing that has enabled the AI revolution that we're living through, which I think we're at the bottom of the first inning in, is one, we had the ability to do cloud computing, so just apply significantly more computational power to old algorithms, and then b, we had dramatically more data.”
Baker states data quantity is the single most predictive element of AI quality, not algorithms or infrastructure.
“The single most predictive element of knowledge about AI quality is the quantity of data used to train the algorithm.”
Baker cites research showing every 10x increase in training data doubles AI quality.
“and it's been very well established in multiple papers from both Google and Microsoft research that for every order of magnitude increase in the data you use to train an algorithm, the quality of the AI doubles.”
https://x.com/GavinSBaker/status/2091542026072338623
post Baker says Atreides internal AI spend will be 100x higher in August 2026 versus March 2026 and still doubling monthly
https://x.com/GavinSBaker/status/2091590204486451277
post Baker estimates his personal AI usage is up 100x and calls Grok Bot another Claude Code moment
https://x.com/GavinSBaker/status/2089379355692527813
post Baker says Grok 4.6 matches Fable 5 Max at 85% discount, 80% cheaper input and 88% cheaper output, Pareto dominant
https://x.com/GavinSBaker/status/2087567239423676519
https://x.com/GavinSBaker/status/2090114970595864991
post Baker would not take under on 250 billion for Anthropic unless they stumble or fail to secure compute
https://x.com/GavinSBaker/status/2089415607162634501
post Baker says future is multi-model with specialized open-source models on customer data working with frontier models behind custom router
https://x.com/GavinSBaker/status/2091887341786833180
post Baker credits Jalapeño as first good ASIC outside TPU and Trainium but says disaggregated GPU plus SRAM will outperform
https://x.com/GavinSBaker/status/2092265540068966504
post Baker says Stripe is growing net revenue and FCF almost 2x faster than Adyen with declining share count
https://x.com/GavinSBaker/status/2090157968377405866
post Baker suggests Anthropic may have shifted from gross to net ARR reporting making 65 billion more conservative than 47 billion
https://x.com/GavinSBaker/status/2090046053395402862