Why This Role Stands Out
This contract-to-hire Senior AWS Site Reliability Engineer role offers a fantastic opportunity to deepen your expertise in cloud infrastructure and automation within a reputable company. You'll thrive here if you possess a strong background in AWS and SRE principles, driving critical improvements in application resilience and operational efficiency. Apply now to contribute to impactful projects and advance your career in a dynamic on-site environment.
Quick Overview
Job Description
Job Title: Senior AWS Site Reliability Engineer (SRE)
Location: Birmingham, Alabama
Type: Contract To Hire
Work Model: Onsite – onsite
Hours: 40.0
Security Clearance:
Overview
Responsibilities
- Implement and improve SRE practices across AWS hosted enterprise applications and services.
- Define and monitor SLIs/SLOs, error budgets, alarms, availability, and reliability metrics.
- Automate infrastructure and operational processes using Python or Java.
- Develop and maintain infrastructure as code using Terraform and AWS native tooling.
- Build and enhance CI/CD pipelines using Jenkins, GitLab, and related technologies.
- Implement monitoring, APM, distributed tracing, dashboards, and alerting using Splunk, SignalFx, OpenTelemetry, and CloudWatch.
- Improve application resilience, failover readiness, recovery automation, and production stability.
- Analyze application and infrastructure performance, identify recurring reliability issues, and implement sustainable improvements.
- Troubleshoot AWS workloads, application performance, data pipelines, networking/DNS, CI/CD, and infrastructure automation issues.
- Support production incidents, releases, platform upgrades, vulnerability remediation, and controlled infrastructure changes.
- Establish performance baselines and help improve monitoring and operational readiness.
- Lead technical discussions and coordinate resolution across application, cloud, DevOps, security, performance, and production support teams.
- 8+ years of hands on AWS engineering experience supporting enterprise production environments.
- Experience across AWS services such as ECS, EC2, RDS, Lambda, Route 53, Step Functions, Redshift, EMR, DynamoDB, S3, and CloudWatch.
- Proven implementation of core Site Reliability Engineering (SRE) practices, including SLIs/SLOs, error budgets, monitoring, alerting, incident response, reliability, and resiliency.
- Strong Terraform / Infrastructure as Code (IaC) experience.
- Experience building and supporting CI/CD pipelines with Jenkins, GitLab, or comparable tooling.
- Strong programming and automation skills using Python or Java.
- Hands on observability experience using Splunk, SignalFx, OpenTelemetry, and CloudWatch across logs, metrics, traces, dashboards, and alerts.
- Experience with APM and distributed tracing in enterprise applications.
- Strong production troubleshooting skills across application, cloud infrastructure, performance, deployment, and reliability issues.
- Experience implementing resilience, recovery, failover, and production stability improvements.
- Ability to troubleshoot complex AWS environments, including unhealthy workloads, failed serverless workflows, database/data pipeline issues, routing/DNS problems, capacity constraints, and automation failures.
- Strong ownership, analytical troubleshooting, communication, and cross team collaboration skills.
System One, and its subsidiaries including Joulé and Mountain Ltd., are leaders in delivering outsourced services and workforce solutions across North America. We help clients get work done more efficiently and economically, without compromising quality. System One not only serves as a valued partner for our clients, but we offer eligible employees health and welfare benefits coverage options including medical, dental, vision, spending accounts, life insurance, voluntary plans, as well as participation in a 401(k) plan.
System One is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, age, national origin, disability, family care or medical leave status, genetic information, veteran status, marital status, or any other characteristic protected by applicable federal, state, or local law.
#M-
#LI-
Similar jobs
- IA
Senior Site Reliability Engineer
NewIO Associates
Menlo Park, CA🇺🇸Hybrid22 hours agoAWSELKAzure+9Technology - BI
SRE lead @ Phoenix AZ 85054 (Hybrid) - Only on W2
NewBURGEON IT SERVICES LLC
Phoenix, AZ🇺🇸Hybrid22 hours agoScrumGoogle CloudJava+1Technology - IN
Infrastructure Automation Engineer
NewInnova
Chandler, AZ🇺🇸$65 - $70/hrHybrid22 hours agoAWSAgileAnsible+4Technology - VC
DevOps Engineer
NewVST Consulting, Inc
Fort Mill, SC🇺🇸Hybrid22 hours agoAWSAzure.NET+2Technology - HS
API Gateway / Platform Engineer
NewHierarch Soft Technologies, Inc.
Wilmington, DE🇺🇸Hybrid22 hours agoDockerMicroservicesSpring+8Technology - BI
TS/SCI - Devops/ Systems Engineer with Security Clearance
NewBailey Information Technology, LLC
Chantilly, VA🇺🇸HybridYesterdayDockerAPI GatewayAWS+5Technology