Site Reliability Engineer, Cloud Infrastructure
In this role, you will be responsible for building, maintaining, and improving the cloud infrastructure that powers Weave's services. You will work with a modern tech stack, including Google Cloud Platform (GCP), Go, Kubernetes, Terraform, Prometheus, Grafana, and Vault. As an engineer, you will be proficient in core tools and languages, capable of completing routine tasks independently, and will use established patterns to create high-quality, maintainable solutions. You will play a key role in ensuring the reliability, scalability, and performance of our platform.
This position will be remote
Reports to: Engineering Manager
What You Will Own
Automate away as much of the day-to-day work as possible.
Design and implement highly available and scalable systems.
Ensure smooth day-to-day operations of Weave’s infrastructure.
Build and evolve tools and standards for automation, scaling, monitoring, and alerting.
Collaborate with product teams to resolve production issues, improve monitoring and leverage cloud services.
Participate in weekly on-call rotation.
<