Haystack
← Back to Jobs
Technology
TR

Site Reliability Engineer (SRE) (Only W2)

Trispark IncScottsdale, AZ🇺🇸United StatesPosted 1 Sept 2026

Quick Overview

Salary
$48/hr
Seniority
Mid Senior
Work mode
On Site
Location
Scottsdale, AZ, United States
Posted
19 hours ago
MongoDBOraclePL/SQLRustSQLSQL ServerAPI GatewayAWSLoad BalancingService MeshSplunkTCP/IPAgileAzureBigQueryConfluenceDNSGenerative AIGoogle CloudGrafanaGraphQLHTTPJavaKubernetesPostgreSQLPrismaPrometheusPythonRedisVault

Job Description

Site Reliability Engineer (SRE)
Duration: 6 Months

Pay Rate: $48/HR
Work Mode: Hybrid
Interview Type:  Interview would be conducted in multiple rounds with different panels focused on Technical Skills & Experience.
Location: 

Scottsdale, Arizona 85260 (Local candidates will be the first preference)

We are seeking a Site Reliability Engineer with strong experience in cloud-native operations, observability, automation, and production support for large-scale enterprise applications.

Skillset required:

·        3-5 years of experience in Site Reliability Engineering, Production Operations, or Platform Engineering supporting large-scale, high-performance applications across hybrid environments (on-premises and cloud).

·        3-5 years of experience developing automation scripts and building Application Performance Management (APM) dashboards to monitor end-to-end transaction journeys.

·        Hands-on programming experience (2+ years) with one or more languages such as Go, Python, Java, or Rust.

·        Working knowledge of relational and NoSQL databases including Oracle, SQL Server, PostgreSQL, MongoDB, Redis, ClickHouse, PL/SQL, or time-series databases.

·        Experience with cloud migration and containerization initiatives using Google Cloud Platform, AWS, Azure, Rancher, OpenShift, or similar platforms.

·        Experience managing containerized applications in Kubernetes environments such as GKE, RKE, or AKS.

·        Strong experience implementing observability solutions using Open Telemetry (OTEL), distributed tracing, monitoring, and incident management.

·        Familiarity with GraphQL frameworks such as Apollo, Prisma, or Hasura.

·        Strong networking fundamentals including TCP/IP, HTTP, DNS, load balancing, and service mesh technologies.

·        Experience participating in 24x7 on-call rotations and meeting incident response SLAs.

Preferred Qualifications:

·        Experience managing highly available, customer-facing platforms with a focus on reliability, automation, and operational excellence.

·        Hands-on experience with monitoring and observability tools such as Splunk, Dynatrace, AppDynamics, Grafana, and Prometheus.

·        Experience with CI/CD and Agile tools such as Rally, Confluence, and related DevOps platforms.

·        Knowledge of in-memory caching technologies, especially Redis.

·        Strong troubleshooting and debugging skills across distributed systems and API gateway architectures.

·        Experience with Google Cloud services including GCS, Cloud SQL, Spanner, and BigQuery.

·        Experience supporting HashiCorp Vault environments.

·        Exposure to Vertex AI, Generative AI, and cloud-based analytics platforms.

Similar jobs