Haystack
← Back to Jobs
Technology
SB

Google Cloud Platform Site Reliability Engineer (SRE) Scottsdale - AZ - Arizona

Sierra Business Solution LLCScottsdale, AZ🇺🇸United StatesPosted Sep 25, 2026

Quick Overview

Seniority
Mid Senior
Work mode
Hybrid
Location
Scottsdale, AZ, United States
Posted
Yesterday
DockerSpringAPI GatewayAWSELKService MeshAnsibleArgoCDAzureBashGitHub ActionsGoogle CloudGrafanaHelmIstioJenkinsKubernetesPowerShellPrometheusPythonTerraform

Job Description

We are looking for a highly experienced Site Reliability Engineer to design, operate, automate, and continuously improve highly reliable, scalable, secure, and observable cloud Google Cloud Platform. The ideal candidate will bring deep hands-on experience in SRE practices, multi-cloud infrastructure, Kubernetes platforms, CICD engineering, infrastructure automation, observability, disaster recovery, incident response, and reliability governance. This role is suited for an engineer with strong experience in SLISLO management, error budget governance, cloud migration, platform modernization, production support, and automation-led toil reduction.

The candidate will be responsible for improving platform availability, deployment reliability, incident response maturity, disaster recovery readiness, observability coverage, and operational efficiency across enterprise cloud environments. The role requires close collaboration with development, security, infrastructure, platform, and business teams to ensure business-critical applications meet defined reliability, performance, compliance, and scalability objectives.

Key Responsibilities

Site Reliability Engineering and Reliability Governance (SLIs, SLOs, error budgets, reliability reviews, and blameless postmortem practices, MTTD, MTTR, recurring incidents)

Kubernetes, Containers, and Platform Engineering (Kubernetes, EKS, AKS, GKE, ECS, and Docker-based platforms)

Infrastructure as Code and Automation (Terraform)

CICD and Release Reliability (GitHub Actions, blue-green, canary, rolling deployments, automated rollback, deployment validation, and automated testing)

Observability, Monitoring, and Logging (Prometheus, Grafana)

Disaster Recovery, High Availability, and Resilience

Security, Compliance, and Cloud Governance

Linux Systems Administration and Production Support

Required Qualifications

10+ years of experience in Site Reliability Engineering, DevOps, cloud infrastructure, Linux administration, production operations, or platform engineering.

Strong hands-on experience designing, operating, and supporting cloud infrastructure across AWS, Azure, and Google Cloud Platform.

Deep experience with Kubernetes platforms such as EKS, AKS, GKE, and containerization using Docker.

Experience building and maintaining CICD pipelines using Jenkins, GitHub Actions.

Strong observability experience with Prometheus, Grafana, ELK Stack, OpenSearch, Log Analytics, Application Insights, and Google Cloud Platform Cloud Monitoring.

Experience with disaster recovery, high availability, backup automation, multi-region failover, and recovery validation.

Hands-on scripting and automation experience using Python, Bash, PowerShell, and Ansible.

Linux systems administration experience across enterprise production environments.

Preferred Qualifications

Experience implementing GitOps using ArgoCD, Helm.

Experience with service mesh and API traffic management using Istio Service Mesh and API Gateway.

Experience supporting regulated enterprise environments in healthcare, financial services, banking, insurance, or similarly controlled domains. The resume includes experience across financial, healthcare, insurance, retail, and banking clients.

Experience with cloud cost optimization.

Experience with security and governance tooling such as GuardDuty, CloudTrail, Kubernetes RBAC, Secrets Manager.

Experience authoring runbooks, DR playbooks, operational procedures, architecture diagrams, and incident response documentation.

Skills: PostgreSQLDigital : DockerDigital : Google CloudDigital : Spring BootDigital : Kubernetes

Experience Required: 8-10

Similar jobs