Haystack
← Back to Jobs
Technology

Data Infrastructure Site Reliability Engineer (SRE)

Prudent Technologies and ConsultingUnited States🇺🇸United StatesPosted 23 Jul 2026

Quick Overview

Work Type
Hybrid
Level
Mid Senior

Job Description

Role: Data Infrastructure Site Reliability Engineer (SRE) – AWS & Big Data Platforms (Apple)

Location: Dallas, Texas (Remote)

Contract to hire after 3 months

Duration: 12 Months

Skills - Site Reliability Engineering, Big Data, AWS, AWS EMR, AWS EKS, AWS MSK, AWS Athena, AWS Glue, Spark, Iceberg, Python, Terraform, Cloudera Hadoop, Agentic Ai, Monitoring Tools

 Job Description:

Required Skills & Experience

Site Reliability Engineering (SRE)

  • Strong SRE mindset with a proven focus on reliability, availability, performance optimization, incident management, and operational excellence.
  • Experience delivering services within defined SLAs and ensuring timely resolution of production issues.
  • Expertise in troubleshooting complex distributed systems and identifying root causes quickly and effectively.

 

  • AWS & Cloud Infrastructure
  • Deep hands-on experience with AWS services, including:
  • EMR
  • EKS
  • MSK
  • Athena
  • Glue
  • IAM
  • Amazon S3
  • VPC
  • AWS networking and security services
  • Strong understanding of cloud-native architectures, scalability, and infrastructure resilience.

Hands-on experience with:

  • Apache Spark
  • Apache Iceberg
  • Big Data platform architecture
  • Performance tuning and optimization

Strong experience with:

  • Terraform
  • Infrastructure as Code (IaC)
  • CI/CD pipeline implementation and automation
  • Platform engineering best practices

 

Skills

AWS
Apache
Apache Spark
Hadoop
Python
Terraform

Similar jobs