Yuri Milyutin’s Post

Nvidia started shipping Vera Rubin this month. Most edge sites can't physically host it. Here's the part that never makes the roadmap slide. Every Nvidia generation makes the rack hotter. A single B300 pulls 1,400 watts. A GB300 NVL72 rack pulls 132 to 140 kW today, liquid cooled, no exceptions. Rubin climbs from there. And inference is the workload driving the volume now. It doesn't want to sit in three mega-campuses. It wants to live near the data and the users, which pushes it out to regional and edge sites. So watch the collision. The silicon is getting denser and hotter. The workload is moving to smaller, distributed locations. And most of the buildings in those locations were spec'd for 5 to 10 kW racks, back when that was a lot. Air cooling quits above 40 kW per rack. That's not a vendor preference. It's physics. You do not turn a 2015 shell into a 130 kW liquid-cooled hall by adding fans. This is the whole case for factory-built modular. You build the power and the cooling for the density up front, in a plant, then ship it to the site. Edge inference at 40 kW per rack and up, cooled properly, live in months instead of years. The roadmap isn't slowing down. Either your real estate catches up to the silicon, or the silicon sits in a crate.

To view or add a comment, sign in

Explore content categories