← All jobs
Site Reliability Engineer, Cloud Infrastructure
Weave - Headquarters (Lehi, UT) · Remote · FullTime · Technology
Apply well, not just fast
Create a free account and upload your resume to get a match score, keyword gaps, a tailored resume, a cover letter and interview prep for this job.
About the role
GCPKubernetesGoTerraformPythonDockerAnsibleCI/CDGitPrometheusGrafanaSRE
In this role, you will be responsible for building, maintaining, and improving the cloud infrastructure that powers Weave's services. You will work with a modern tech stack, including Google Cloud Platform (GCP), Go, Kubernetes, Terraform, Prometheus, Grafana, and Vault. As an engineer, you will be proficient in core tools and languages, capable of completing routine tasks independently, and will use established patterns to create high-quality, maintainable solutions. You will play a key role in ensuring the reliability, scalability, and performance of our platform.
- This position will be remote
- Reports to: Engineering Manager
WHAT YOU WILL OWN
- Automate away as much of the day-to-day work as possible.
- Design and implement highly available and scalable systems.
- Ensure smooth day-to-day operations of Weave’s infrastructure.
- Build and evolve tools and standards for automation, scaling, monitoring, and alerting.
- Collaborate with product teams to resolve production issues, improve monitoring and leverage cloud services.
- Participate in weekly on-call rotation.
WHAT YOU WILL NEED TO ACCOMPLISH THE JOB
- Proficiency with at least one cloud platform is required, with GCP services (e.g., GKE, Compute Engine, VPC) being a plus.
- Solid understanding of containerization technologies such as Kubernetes and Docker.
- Experience with automation tools such as Puppet, Salt, Ansible, and Terraform.
- Experience writing automation using Go, Python, etc.
- Experience designing highly available and scalable systems.
- Proficient with version control systems (e.g., Git) and CI/CD concepts.
- Strong problem-solving skills and the ability to troubleshoot complex issues systematically.
WHAT WILL MAKE US LOVE YOU
- A passion for Infrastructure as Code, constantly seeking opportunities to automate, optimize, and manage infrastructure through code.
- Deep expertise in Kubernetes, including cluster design, deployment, and ongoing management for large-scale applications.
- Experience with advanced GCP services and architectures.
- A strong sense of ownership and accountability for the systems you build and maintain.
- Managing infrastructure and applications using IaC, GitOps and ArgoCD.
Employment with Weave is contingent upon the successful completion of a background check, conducted in accordance with applicable laws.
At Weave, we use Artificial Intelligence (AI) tools to help us work more efficiently and create a smoother candidate experience. AI may assist with things like writing job descriptions, scheduling interviews, or reviewing applications against job-related criteria. For additional information, please review the External AI Policy Statement available on our Careers page.
Weave is an equal opportunity employer that is committed to fostering an inclusive workplace where all individuals are valued and supported. We welcome anyone who is hungry to learn, problem-solve and progress regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, or other applicable legally protected characteristics. If you have a disability or special need that requires accommodation, please let us know.
Beware of recruitment fraud. All official correspondence will occur through Weave branded email. We will never ask you to share bank account information, cash a check from us, or purchase software or equipment as part of your interview or hiring process.