Haystack
← Back to Jobs
Technology
IV

Site Reliability Engineer

Indus ValleyRiverwoods, IL🇺🇸United StatesPosted Sep 18, 2026

Quick Overview

Seniority
Mid Senior
Work mode
On Site
Location
Riverwoods, IL, United States
Posted
18 hours ago
MySQLSQLShellAWSELKAgileAnsibleDatadogGrafanaHadoopJavaJenkinsJiraKafkaKibanaKubernetesPython

Job Description

Title: Site Reliability EngineerLocation: Riverwoods, IL (hybrid - 2 to 3 days/week onsite)Duration: 12+ Months. C2H
*Look for someone with monitoring engineering exp (Dynatrace, Datadog, etc.)
Must have:
  • Professional experience as a Site Reliability Engineer (SRE)
  • Experience in performance testing, Ability to translate functional and non-functional requirements into appropriate NFT Automation tests.
  • Experience of AWS Cloud Application (Must for sure)
  • Experience Linux, AWS Cloud(Must) and Prem deployments .
  • Good experience in Systems Observability and APM tools, preferably Datadog
  • Experience in dashboarding tools such as Grafana and Kibana
  • Strong ability to track and contribute to technical discussions around application integration and high-availability, resilience and observability.
  • Expertise in one or more programming languages: Python, shell scripting (Unix/Linux), Java

Nice to have:
  • Hands-on experience on SNOW.
  • Experience in container technology (OpenShift, Kubernetes)
  • Strong JIRA knowledge
  • Basic understating of Release Management.
  • Experience in CI/CD pipelines preferably Jenkins expertise.

JD:
  • Experience 12to 15 years
  • Professional experience as a Site Reliability Engineer (SRE)
  • Software development hands on engineer with excellent understanding of SDLC Application delivery.
  • Ability to translate functional and non-functional requirements into appropriate NFT Automation tests.
  • Experience with DevOps, CI/CD tools.
  • Good experience of Linux, AWS Cloud and on Prem deployments
  • Good experience in Systems Observability and APM tools, preferably Datadog
  • Strong ability to track and contribute to technical discussions around application integration and high-availability, resilience and observability.
  • Strong JIRA knowledge

Responsibilities
  • Partner with Application Development teams to build resiliency for Payment application.
  • Partner with our Application Develop teams to implement service level objectives.
  • Partner with our Application Development teams and other SREs to build out end to end observability.
  • Implement monitoring, alerting and dashboards needed for our apps.
  • Automated operational processes.
  • Help to develop our capacity management and performance management tools.
  • Help to define the DR plan needed for our critical apps.
  • Help to develop a chaos testing process.
  • Participate in an on-call rotation and support production Incidents
  • SRE Skillsets Expectations from Discover for Pricing & Settlements:
  • Good understanding of hybrid infrastructure
  • Expertise with AWS
  • Expertise in one or more general purpose programming languages: Python, Go, shell scripting (Unix/Linux), Java
  • Experience in CI/CD pipelines preferably Jenkins expertise.
  • Experience in container technology (OpenShift, Kubernetes)
  • Expertise in automation tools experience (preferably Ansible).
  • Expertise in observability tools including APM (Datadog), synthetic monitoring and log aggregation (Elk)
  • Experience in dashboarding tools such as Grafana and Kibana
  • Understating of Agile concepts and experience in JIRA
  • Basic understating of Release Management
  • Hands-on experience on SNOW
  • SRE Skillsets Expectations from Discover for Data Platform:
  • Expertise in Message Broker (preferably Rabbit MQ, Kafka)
  • Expertise on Hadoop, spark commands JSON formatting

Skill
  • Expertise in AWS-Lambda Services 4/5
  • Strong programming skills (Java, Python, Shell Optional Java Script) - 4/5
  • Proficiency in Database concepts, Strong knowledge of SQL and experience with MySQL 4/5
  • APM tools - Datadog 4/5
  • Hands-on experience on SNOW

Similar jobs