On the record about
2 people · 3 quotes · 13 Oct 2024 to 3 Oct 2025
2 of 2 lanes rest on fewer than 5 quotes and are marked thin. Offsets are days from the middle first-quote date, 13 Oct 2024 — a date, and nothing else. It is not a claim about who reached a view first.
Gurley explains NVIDIA's networking, NVLink, and CUDA advantages emerge specifically in the largest system deployments.
“That's when the networking piece thrives. That's where NVLink thrives. That's where CUDA really comes alive in the biggest systems that are out there.”
Patel explains o1's thinking time creates memory bandwidth issues that prevent batching users at high levels.
“But if you batch higher, k b cache is not just a memory capacity issue, it's also a memory bandwidth issue.”
Patel explains GB200's 72-GPU configuration creates reliability challenges compared to 8-GPU systems due to higher failure probability.
“If the reliability of each GPU is the same, then when a single GPU fails in 72 GPUs, you have a much higher chance of something failing, right?”