Principal Engineer, Rack-Scale GPU Architecture
Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, naturally like to think big and bold, and are energized by the fast-paced environment of a true industry disruptor, you’ll find your place here. We value winning together—while learning, having fun, and making a profound difference for the dreamers and builders in the world.
We are looking for a Principal Engineer to own how rack-scale GPU systems are architected, orchestrated, and operated for inference at DigitalOcean.
DigitalOcean is the Inference Cloud. The unit of GPU capacity is changing underneath the entire industry: for a decade the server was the boundary, and now it is the rack. GB300 NVL72 and the Vera Rubin generation behind it present 72 or more GPUs inside a single coherent NVLink domain, which breaks most of the assumptions our scheduling, networking, failure-handling, and capacity models were built on. Getting this right determines whether we can serve trillion-parameter models with the economics our customers expect.
This role owns that transition. You will define the reference architecture for rack-scale inference at DigitalOcean—how NVLink domains are carved and allocated, how t