Google Cloud Platform Site Reliability Engineer (SRE) Scottsdale - AZ - Arizona
Quick Overview
Job Description
We are looking for a highly experienced Site Reliability Engineer to design, operate, automate, and continuously improve highly reliable, scalable, secure, and observable cloud Google Cloud Platform. The ideal candidate will bring deep hands-on experience in SRE practices, multi-cloud infrastructure, Kubernetes platforms, CICD engineering, infrastructure automation, observability, disaster recovery, incident response, and reliability governance. This role is suited for an engineer with strong experience in SLISLO management, error budget governance, cloud migration, platform modernization, production support, and automation-led toil reduction.
The candidate will be responsible for improving platform availability, deployment reliability, incident response maturity, disaster recovery readiness, observability coverage, and operational efficiency across enterprise cloud environments. The role requires close collaboration with development, security, infrastructure, platform, and business teams to ensure business-critical applications meet defined reliability, performance, compliance, and scalability objectives.
Key Responsibilities
Site Reliability Engineering and Reliability Governance (SLIs, SLOs, error budgets, reliability reviews, and blameless postmortem practices, MTTD, MTTR, recurring incidents)
Kubernetes, Containers, and Platform Engineering (Kubernetes, EKS, AKS, GKE, ECS, and Docker-based platforms)
Infrastructure as Code and Automation (Terraform)
CICD and Release Reliability (GitHub Actions, blue-green, canary, rolling deployments, automated rollback, deployment validation, and automated testing)
Observability, Monitoring, and Logging (Prometheus, Grafana)
Disaster Recovery, High Availability, and Resilience
Security, Compliance, and Cloud Governance
Linux Systems Administration and Production Support
Required Qualifications
10+ years of experience in Site Reliability Engineering, DevOps, cloud infrastructure, Linux administration, production operations, or platform engineering.
Strong hands-on experience designing, operating, and supporting cloud infrastructure across AWS, Azure, and Google Cloud Platform.
Deep experience with Kubernetes platforms such as EKS, AKS, GKE, and containerization using Docker.
Experience building and maintaining CICD pipelines using Jenkins, GitHub Actions.
Strong observability experience with Prometheus, Grafana, ELK Stack, OpenSearch, Log Analytics, Application Insights, and Google Cloud Platform Cloud Monitoring.
Experience with disaster recovery, high availability, backup automation, multi-region failover, and recovery validation.
Hands-on scripting and automation experience using Python, Bash, PowerShell, and Ansible.
Linux systems administration experience across enterprise production environments.
Preferred Qualifications
Experience implementing GitOps using ArgoCD, Helm.
Experience with service mesh and API traffic management using Istio Service Mesh and API Gateway.
Experience supporting regulated enterprise environments in healthcare, financial services, banking, insurance, or similarly controlled domains. The resume includes experience across financial, healthcare, insurance, retail, and banking clients.
Experience with cloud cost optimization.
Experience with security and governance tooling such as GuardDuty, CloudTrail, Kubernetes RBAC, Secrets Manager.
Experience authoring runbooks, DR playbooks, operational procedures, architecture diagrams, and incident response documentation.
Skills: PostgreSQLDigital : DockerDigital : Google CloudDigital : Spring BootDigital : Kubernetes
Experience Required: 8-10
Similar jobs
- OP
Databricks Platform and DevSecOps Engineer
NewOpenDataJobs
United States🇺🇸Remote18 hours agoAWSAzureBash+9 - SO
Senior DevSecOps Engineer
NewSystem One
Knoxville, TN🇺🇸HybridYesterdayEngineering - MR
Ai Platform Engineer
NewMotion Recruitment Partners, LLC
Waltham, MA🇺🇸$127k - $220k/yrOn-siteYesterdayMicroservicesMLOpsMachine Learning+3Technology - RH
Agent Platform Engineer
NewRobert Half
Los Angeles, CA🇺🇸$200k - $300k/yrOn-siteYesterdaySQLMachine LearningPython+2Technology - CR
Cloud Platform Engineer
NewCredence
McLean, Virginia🇺🇸$100k - $180k/yrHybridYesterdayDockerShellAWS+18Technology - SO
Senior DevSecOps Engineer
NewSystem One
Birmingham, AL🇺🇸HybridYesterdayEngineering