Haystack
← Back to Jobs
Technology
II

Site Reliability Engineer

Icon International Group LLCScottsdale, AZ🇺🇸United StatesPosted 9 Sept 2026

Why This Role Stands Out

This hybrid Site Reliability Engineer role offers a fantastic opportunity to contribute to large-scale enterprise applications, honing your skills in cloud-native operations and automation. You'll thrive here if you're a mid-senior professional with a passion for building robust, scalable systems and possess strong programming and cloud experience. Apply today to join a dynamic team and advance your career!

Quick Overview

Seniority
Mid Senior
Work mode
Hybrid
Location
Scottsdale, AZ, United States
Posted
15 hours ago
MongoDBOraclePL/SQLRustSQLSQL ServerAPI GatewayAWSLoad BalancingService MeshSplunkTCP/IPAgileAzureBigQueryConfluenceDNSGenerative AIGoogle CloudGrafanaGraphQLHTTPJavaKubernetesPostgreSQLPrismaPrometheusPythonRedisVault

Job Description

Position: Site Reliability Engineer (SRE)

Duration: Contract to hire 
Work Mode: Hybrid
Interview Type:  Interview would be conducted in multiple rounds with different panels focused on Technical Skills & Experience.

Location: Scottsdale, Arizona 

 

We are seeking a Site Reliability Engineer with strong experience in cloud-native operations, observability, automation, and production support for large-scale enterprise applications.

Skillset required:

·         3-5 years of experience in Site Reliability Engineering, Production Operations, or Platform Engineering supporting large-scale, high-performance applications across hybrid environments (on-premises and cloud).

·         3-5 years of experience developing automation scripts and building Application Performance Management (APM) dashboards to monitor end-to-end transaction journeys.

·         Hands-on programming experience (2+ years) with one or more languages such as Go, Python, Java, or Rust.

·         Working knowledge of relational and NoSQL databases including Oracle, SQL Server, PostgreSQL, MongoDB, Redis, ClickHouse, PL/SQL, or time-series databases.

·         Experience with cloud migration and containerization initiatives using Google Cloud Platform, AWS, Azure, Rancher, OpenShift, or similar platforms.

·         Experience managing containerized applications in Kubernetes environments such as GKE, RKE, or AKS.

·         Strong experience implementing observability solutions using Open Telemetry (OTEL), distributed tracing, monitoring, and incident management.

·         Familiarity with GraphQL frameworks such as Apollo, Prisma, or Hasura.

·         Strong networking fundamentals including TCP/IP, HTTP, DNS, load balancing, and service mesh technologies.

·         Experience participating in 24x7 on-call rotations and meeting incident response SLAs.

Preferred Qualifications:

·         Experience managing highly available, customer-facing platforms with a focus on reliability, automation, and operational excellence.

·         Hands-on experience with monitoring and observability tools such as Splunk, Dynatrace, AppDynamics, Grafana, and Prometheus.

·         Experience with CI/CD and Agile tools such as Rally, Confluence, and related DevOps platforms.

·         Knowledge of in-memory caching technologies, especially Redis.

·         Strong troubleshooting and debugging skills across distributed systems and API gateway architectures.

·         Experience with Google Cloud services including GCS, Cloud SQL, Spanner, and BigQuery.

·         Experience supporting HashiCorp Vault environments.

·         Exposure to Vertex AI, Generative AI, and cloud-based analytics platforms.

Similar jobs