Why This Role Stands Out
This remote Site Reliability Engineer role offers significant growth by empowering you to build and scale cutting-edge cloud infrastructure with a modern tech stack. You'll thrive here if you're a driven engineer eager to automate, innovate, and ensure the reliability of a dynamic platform. Apply now to join a forward-thinking company and make a real impact.
Quick Overview
Job Description
In this role, you will be responsible for building, maintaining, and improving the cloud infrastructure that powers Weave's services. You will work with a modern tech stack, including Google Cloud Platform (GCP), Go, Kubernetes, Terraform, Prometheus, Grafana, and Vault. As an engineer, you will be proficient in core tools and languages, capable of completing routine tasks independently, and will use established patterns to create high-quality, maintainable solutions. You will play a key role in ensuring the reliability, scalability, and performance of our platform.
This position will be remote
Reports to: Engineering Manager
What You Will Own
Automate away as much of the day-to-day work as possible.
Design and implement highly available and scalable systems.
Ensure smooth day-to-day operations of Weave’s infrastructure.
Build and evolve tools and standards for automation, scaling, monitoring, and alerting.
Collaborate with product teams to resolve production issues, improve monitoring and leverage cloud services.
Participate in weekly on-call rotation.
What You Will Need to Accomplish the Job
Proficiency with at least one cloud platform is required, with GCP services (e.g., GKE, Compute Engine, VPC) being a plus.
Solid understanding of containerization technologies such as Kubernetes and Docker.
Experience with automation tools such as Puppet, Salt, Ansible, and Terraform.
Experience writing automation using Go, Python, etc.
Experience designing highly available and scalable systems.
Proficient with version control systems (e.g., Git) and CI/CD concepts.
Strong problem-solving skills and the ability to troubleshoot complex issues systematically.
What Will Make Us Love You
A passion for Infrastructure as Code, constantly seeking opportunities to automate, optimize, and manage infrastructure through code.
Deep expertise in Kubernetes, including cluster design, deployment, and ongoing management for large-scale applications.
Experience with advanced GCP services and architectures.
A strong sense of ownership and accountability for the systems you build and maintain.
Managing infrastructure and applications using IaC, GitOps and ArgoCD.
Employment with Weave is contingent upon the successful completion of a background check, conducted in accordance with applicable laws.
At Weave, we use Artificial Intelligence (AI) tools to help us work more efficiently and create a smoother candidate experience. AI may assist with things like writing job descriptions, scheduling interviews, or reviewing applications against job-related criteria. For additional information, please review the External AI Policy Statement available on our Careers page.
Weave is an equal opportunity employer that is committed to fostering an inclusive workplace where all individuals are valued and supported. We welcome anyone who is hungry to learn, problem-solve and progress regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity, veteran status, or other applicable legally protected characteristics. If you have a disability or special need that requires accommodation, please let us know.
Beware of recruitment fraud. All official correspondence will occur through Weave branded email. We will never ask you to share bank account information, cash a check from us, or purchase software or equipment as part of your interview or hiring process.
Similar jobs
- ON
Site Reliability Engineer II
NewAuto ApplyOnapsis
Dallas🇺🇸Hybrid11 hours agoOracleAWSTDD+10Technology - NE
Site Reliability Engineer
NewAuto ApplyNebius
Remote - United States🇺🇸$130k - $180k/yrRemote4 hours agoBashPythonTechnology - EL
Senior DevOps Engineer (NOAA badge required)
NewAuto ApplyElement84
Alexandria HQ (remote)🇺🇸Remote11 hours agoDockerDynamoDBSQL+20Technology - CM
Senior AI Platform Engineer
NewAuto ApplyCode Metal
Boston Hub🇺🇸Hybrid2 hours agoAPI GatewayMLflowSalesforce+7Technology - CL
Platform Engineer (Kubernetes), Mid-Level
NewAuto ApplyClera
San Francisco🇺🇸On-site12 hours agoService MeshArgoCDGrafana+4Technology - CO
Senior Site Reliability Engineer
NewAuto ApplyCoalition, Inc.
Any location🇺🇸Hybrid6 hours agoMicroservicesAWSCapacity Planning+7Technology - CI
DevSecOps Engineer
NewAuto ApplyCHAOS Industries
Washington🇺🇸Hybrid7 hours agoDockerAWSOWASP+14Engineering - SA
Senior Infrastructure & Reliability Engineer
NewAuto ApplyStuut Ai
New York City🇺🇸Hybrid8 hours agoSOC 2AuditingERP+2Technology - CA
Associate Site Reliability Engineer/Site Reliability Engineer
NewAuto ApplyC3 AI
Redwood City🇺🇸Hybrid13 hours agoGCPAWSAnsible+8Technology - AI
Contractor: DevOps Engineer
NewAuto ApplyAbacus Insights
United States🇺🇸Hybrid13 hours agoDockerAWSSplunk+10Technology - ED
DevOps Engineer III
NewAuto ApplyEnable Dental
Austin, Texas🇺🇸Remote13 hours agoDockerAWSEncryption+17Technology - TH
DevOps Engineer
NewAuto ApplyTheIncLab
Colorado Springs, Colorado🇺🇸Hybrid7 hours agoDockerShellAWS+22Technology