Senior Site Reliability Engineer
AI companies need inference that’s fast, reliable, and economical at scale. Parasail delivers it. We’re building an enterprise-grade inference cloud for open-weight models where customers pay for the tokens they use, and we handle everything required to serve them.
Behind one OpenAI-compatible API, we pool GPU capacity from providers around the world and continuously optimize where and how workloads run. That means turning a changing mix of hardware, networks, and infrastructure into a service customers can trust.
We’ve raised a $32 million Series A, and we’re scaling beyond trillions of tokens a day. You’ll have the ownership and reach to shape how we get there.
The Role
At Parasail, reliability is an engineering problem that spans the entire stack. A GPU fails. A provider goes down. Traffic spikes. Customers still expect their inference to work.
We’re hiring Site Reliability Engineers to build the systems that make that possible. You’ll own infrastructure across our global GPU fleet, write software that automates operations, and make the platform better at detecting, surviving, and recovering from failures.
You’ll work directly with infrastructure, platform, and inference engineers in a flat organization. We welcome SREs, software engineers, platform engi