Haystack
← Back to Jobs
Technology
IP

Site Reliability Engineer (SRE) – GPS & YAVA Platform

iTek People, Inc.Hartford, CT🇺🇸United StatesPosted Oct 6, 2026

Quick Overview

Seniority
Mid Senior
Work mode
On Site
Location
Hartford, CT, United States
Posted
17 hours ago
Load BalancingAzureDNSGoogle CloudJavaPowerShellPythonWAF

Job Description

Job Title: Site Reliability Engineer (SRE) – GPS & YAVA Platform

Job Type: Contract-to-Hire (C2H)
Location: Hartford, CT or Wellesley, MA
Work Model: Onsite

Job Description:

We are looking for a hands-on Site Reliability Engineer (SRE) to support the GPS & YAVA platform. The role focuses on production reliability, proactive monitoring, automation, incident prevention, change validation, and resilience testing.

Required Skills:

  • Strong experience in SRE, Production Engineering, DevOps, Platform Engineering, or Infrastructure Engineering.

  • Experience with monitoring, logging, tracing, dashboards, alerting, and synthetic monitoring.

  • Strong troubleshooting skills across distributed applications and infrastructure.

  • Experience with DNS, networking, load balancing, TLS/certificates, WAF, API gateways, and cloud services.

  • Programming/scripting experience with Python, Go, PowerShell, Java, or similar.

  • Experience with incident response, RCA, postmortems, change validation, and automation.

  • Experience with resilience, capacity, failover, recovery, or performance testing.

Preferred Skills:

  • Salesforce CRM / CTI integrations

  • Contact Center / Five9 / Telephony

  • IBM MQ

  • DB2 / Mainframe

  • F5

  • Imperva

  • Azure / Google Cloud

  • Chaos Engineering / Performance Engineering

Responsibilities:

  • Monitor and improve GPS & YAVA platform reliability.

  • Build proactive monitoring, alerts, dashboards, and synthetic checks.

  • Validate high-risk application and infrastructure changes.

  • Troubleshoot production incidents and perform root cause analysis.

  • Develop automation for health checks, configuration validation, DNS/routing, certificates, and post-change verification.

  • Support resilience, capacity, failover, and recovery testing.

  • Work closely with application, network, security, database, cloud, telephony, MQ/mainframe, and vendor teams.

Important: This is a C2H opportunity and requires onsite work in Hartford, CT or Wellesley, MA.

Interested candidates: Please share your updated resume and contact details.

Similar jobs