Skip to main content

Greg Steinbrecher

May 6, 2026

12:47
We can't have that happen, so we have to do things differently to make that work.
12:51
Maybe...
12:51
Yeah, so the, the, the very simple math here is, like, you can basically assume that if failures are independent and you double the size of your system, you're going to have half the time between failures, right? Your meantime to failure goes down by half.
13:04
The important thing to think about here for the network is for every GPU, we have tens if not hundreds of network components.
13:12
So even just, like, say you've got one GPU connected to one network adapter.
13:17
In that network adapter, if it has an optical transceiver in it, maybe you'll have four lasers.
13:22
On the other end of that transceiver, you'll have another four lasers.

7 MINS LATER

20:30
Hmm

We value your privacy

We use cookies to understand how you use our platform and to improve your experience. Click “Accept All” to consent, or “Decline non-essential” to opt out of non-essential cookies. Read our Privacy Policy.