← Back to Jobs
Technology
Site Reliability Engineer (SRE) with incident management
Next Gen IT IncKalamazoo, MI🇺🇸United StatesPosted 6 Aug 2026
Quick Overview
Work Type
Hybrid
Level
Mid Senior
Job Description
Sr. Site Reliability Engineer (SRE)
Location: Kalamazoo, MI (3 days a week hybrid)
Exp: $ 13+yrs
Any Visa type is okay
Relocation is also okay
Summary:
The Site Reliability Engineer (SRE) is responsible for improving the reliability, availability, performance, and operability of client’ supported software systems. This role combines software engineering and IT operations to automate operational work, monitor system performance, and reduce toil. The SRE establishes and manages monitoring, alerting, incident response, and problem management practices to ensure applications remain available and performant during updates and failures. The role partners with engineering, architecture, and product teams to define reliability standards and production readiness requirements. SRE is a practical implementation of DevOps focused on maintaining software quality in fast-paced development environments.
Job Description:
Required:
- Strong grounding in SRE/DevOps practices: incident management, blameless postmortems, SLOs/SLIs, error budgets, production readiness.
- Experience building/operating monitoring and alerting, and using logs/metrics to diagnose issues.
- Automation/scripting skills (e.g., Python, PowerShell, Bash) and ability to reduce manual operational work.
- Strong understanding of cloud-based platforms such as Azure DataBricks + Unity Catalog, AWS S3 and RDS.
- Strong experience in ETL / ELT work.
- Understanding of CI/CD concepts, safe deployment patterns, rollback strategies, and change risk controls.
Preferred
- Experience with cloud environments and infrastructure-as-code.
- Experience with large datasets (Multi-million row datasets).
- Experience with container orchestration and modern runtime platforms (where applicable).
- Experience building dashboards and reliability reporting for executives and delivery teams.
Skills
AWS
ETL
Azure
Bash
Databricks
PowerShell
Python
SAFe
Unity
Similar jobs
DevOps I – Linux, AWS, Release Management
Partner's Consulting, Inc. · Englewood, United States
9 minutes agoSite Reliability Engineer
GovCIO · United States
14 minutes ago$230k - $250k/yrAMS / Site Reliability Engineer (SRE) – Telematics & Connected Car
Info Way Solutions · Irvine, United States
15 minutes agoSr. Site Reliability Engineer (SRE)
Next Gen IT Inc · Kalamazoo, United States
16 minutes agoDevOps Engineer
QUANTUM TECHNOLOGIES LLC · Dallas, United States
26 minutes ago€73/hrLLM DevOps/Inference Engineer (W2 Only)
Zuplon · United States
26 minutes ago