Dylan Patel · Open Compute Project · 23 October 2025
“the standard unit for an inference deployment being hundreds of GPUs instead of a single node. And then there's all these different things about traffic.”
Verbatim excerpt with a timestamp. The full recording is at the source; we link out and do not host it. Everything Dylan Patel is on record saying.