Why This Role Stands Out
This hybrid Site Reliability Engineer role offers a fantastic opportunity to contribute to large-scale enterprise applications, honing your skills in cloud-native operations and automation. You'll thrive here if you're a mid-senior professional with a passion for building robust, scalable systems and possess strong programming and cloud experience. Apply today to join a dynamic team and advance your career!
Quick Overview
Job Description
Position: Site Reliability Engineer (SRE)
Duration: Contract to hire
Work Mode: Hybrid
Interview Type: Interview would be conducted in multiple rounds with different panels focused on Technical Skills & Experience.
Location: Scottsdale, Arizona
We are seeking a Site Reliability Engineer with strong experience in cloud-native operations, observability, automation, and production support for large-scale enterprise applications.
Skillset required:
· 3-5 years of experience in Site Reliability Engineering, Production Operations, or Platform Engineering supporting large-scale, high-performance applications across hybrid environments (on-premises and cloud).
· 3-5 years of experience developing automation scripts and building Application Performance Management (APM) dashboards to monitor end-to-end transaction journeys.
· Hands-on programming experience (2+ years) with one or more languages such as Go, Python, Java, or Rust.
· Working knowledge of relational and NoSQL databases including Oracle, SQL Server, PostgreSQL, MongoDB, Redis, ClickHouse, PL/SQL, or time-series databases.
· Experience with cloud migration and containerization initiatives using Google Cloud Platform, AWS, Azure, Rancher, OpenShift, or similar platforms.
· Experience managing containerized applications in Kubernetes environments such as GKE, RKE, or AKS.
· Strong experience implementing observability solutions using Open Telemetry (OTEL), distributed tracing, monitoring, and incident management.
· Familiarity with GraphQL frameworks such as Apollo, Prisma, or Hasura.
· Strong networking fundamentals including TCP/IP, HTTP, DNS, load balancing, and service mesh technologies.
· Experience participating in 24x7 on-call rotations and meeting incident response SLAs.
Preferred Qualifications:
· Experience managing highly available, customer-facing platforms with a focus on reliability, automation, and operational excellence.
· Hands-on experience with monitoring and observability tools such as Splunk, Dynatrace, AppDynamics, Grafana, and Prometheus.
· Experience with CI/CD and Agile tools such as Rally, Confluence, and related DevOps platforms.
· Knowledge of in-memory caching technologies, especially Redis.
· Strong troubleshooting and debugging skills across distributed systems and API gateway architectures.
· Experience with Google Cloud services including GCS, Cloud SQL, Spanner, and BigQuery.
· Experience supporting HashiCorp Vault environments.
· Exposure to Vertex AI, Generative AI, and cloud-based analytics platforms.
Similar jobs
- JM
Site Reliability Engineer III
NewJ.P. Morgan
Houston, Texas🇺🇸On-site3 minutes agoSpringSpring BootJava+2Technology - ST
Senior Site Reliability Engineer
NewStord
United States🇺🇸Hybrid3 minutes agoDockerGCPCloudflare+13Technology - JM
Site Reliability Engineer III - AWS, Java and Kubernetes
NewJ.P. Morgan
Chicago, Illinois🇺🇸On-site3 minutes agoDockerSplunkAnsible+7Technology - OR
Senior Site Reliability Engineer - Oracle Health (US CITIZEN)
NewOracle
United States🇺🇸$81.1k - $187k/yrHybrid3 minutes agoGCPOracleSQL+14Technology - VA
Senior Site Reliability Champion
NewVanguard
Dallas, Texas🇺🇸Hybrid3 minutes agoSplunkPythonTechnology - BA
Staff Engineer, Site Reliability
NewBabylist
United States🇺🇸$226.7k - $272.0k/yrRemote32 minutes agoMySQLRubyAWS+9Technology