Haystack
← Back to Jobs
Technology
ST

Site Reliability Engineer (SRE)

StatusNeo Inc.United States🇺🇸United StatesPosted Sep 24, 2026

Quick Overview

Seniority
Mid Senior
Work mode
Hybrid
Location
United States
Posted
23 hours ago
DockerSQLAWSSplunkAzureBashCircleCICloudFormationDatadogGitHub ActionsGitLab CIGoogle CloudGrafanaHIPAAHelmJenkinsKubernetesPowerShellPrometheusPythonTerraformVault

Job Description

Site Reliability Engineer (SRE)  

We are seeking a highly technical Site Reliability Engineer to build, operate, automate, and scale cloud-native production platforms supporting mission-critical applications in regulated healthcare environments. 

Core Responsibilities 

  • Own production reliability, availability, scalability and performance across cloud infrastructure and application services. 

  • Define and implement SLOs, SLIs, error budgets, alerting strategies and incident response processes. 

  • Partner with software engineering teams to embed reliability, resiliency and operational readiness into the SDLC. 

  • Lead root cause analysis, incident management and post-mortem activities. 

  • Automate infrastructure provisioning, deployments and operational workflows. 

Required Technical Skills 

  • Cloud Platforms: AWS, Azure or Google Cloud Platform. 

  • Containers & Orchestration: Docker, Kubernetes, Helm. 

  • Infrastructure as Code: Terraform or CloudFormation. 

  • CI/CD: Jenkins, GitHub Actions, GitLab CI/CD or CircleCI. 

  • Databases: Operational support, backup, recovery and performance tuning of SQL/NoSQL databases. 

  • Scripting: Python, Bash or PowerShell. 

Observability & Reliability Engineering 

  • Hands-on experience with Datadog, Prometheus, Grafana, CloudWatch, Splunk or Dynatrace. 

  • Build dashboards, logging pipelines, tracing solutions and proactive monitoring frameworks. 

  • Experience diagnosing latency, throughput, capacity and availability issues in distributed systems. 

Security & Compliance 

  • Knowledge of HIPAA or equivalent healthcare regulations. 

  • Secrets management using Vault, AWS Secrets Manager or Azure Key Vault. 

  • Strong understanding of IAM, RBAC, network security and least-privilege access models. 

Qualifications 

  • 3-6+ years of experience in SRE, DevOps or Production Engineering. 

  • Experience supporting highly available customer-facing production services. 

  • Strong troubleshooting and incident management capabilities. 

 

Similar jobs