Site Reliability Engineer
Full Job Description
About the Role
Weave is seeking an experienced Site Reliability Engineer to architect and maintain the robust cloud infrastructure supporting our critical healthcare services. You will focus on building highly available, scalable systems using Google Cloud Platform (GCP), Go, Kubernetes, Terraform, Prometheus, Grafana, and Vault.
Key Responsibilities
- Automate day-to-day operational tasks to maximize efficiency.
- Design and implement fault-tolerant, high-performance infrastructure systems.
- Spearhead the development of tools for automation, scaling, monitoring, and alerting.
- Collaborate cross-functionally with product teams to resolve production issues and optimize cloud services.
- Participate in our weekly on-call rotation to ensure system reliability.
Requirements
We require proficiency with at least one major cloud platform, deep expertise in containerization (Kubernetes/Docker), and extensive experience with automation tools like Puppet, Salt, Ansible, and Terraform. You must be comfortable writing automation scripts in Go or Python.
Company
Weave
Weave is a technology company dedicated to empowering small and medium-sized healthcare practices through advanced AI solutions.Our platform revolutionizes patient engagement, communication, and payme...