Cloudzero logo
Posted 2mo ago•Boston, MA

Senior Platform Engineer

SeniorOn-site (Boston)Salary undisclosed
Required Skills
PythonAWSGCPAzureKubernetesTerraformKafkaLLMs
Job Description

About the Role

CloudZero is growing fast. Our customer base is expanding, the data challenges we're solving are getting more complex, and the platform is scaling to match. We're standing up real-time ingestion on Kafka right now, spanning several engineering teams, and it's the most operationally demanding thing we've built. Nobody owns the reliability of that path end to end today. That's the first thing you'd own. As a Senior Site Reliability Engineer you'll be a force multiplier for our engineering organization, owning the reliability, performance, and observability of the systems every team depends on, and empowering teams to ship features that help customers understand and optimize their cloud spend.

This is real infrastructure work at real scale, not a ticket-closing role and not a console-clicking job. CloudZero processes billions of events daily across AWS, Azure, and GCP. Our customers rely on real-time, accurate cost data to make business-critical decisions, and any instability in our system impacts their planning. Built entirely on a unique serverless architecture with no EC2s and no containers, our platform demands infrastructure that scales gracefully, fails predictably, and recovers automatically. There are no Kubernetes clusters or broker fleets to tune here, so the reliability work is engineering, not firefighting. On-call is light: Nimbus carries a weekly rotation for shared infrastru

Similar Openings in DevOps & Cloud

View all in category➔