Bg2 Pod
“with Blackwell, not only is it way, way, way faster, anywhere from 10 to 15 times on really large models for inference because they've optimized it for very large language models.”
Patel says Blackwell delivers five times performance TCO improvement in single year, accelerating LLM cost decline.
“At least that's what Blackwell is, and we'll see what Ruben does. But, you know, five x plus in a single year for performance TCO is an insane pace.”