Quick Overview
Seniority
Mid Senior
Work mode
On Site
Location
Naperville, IL, United States
Posted
Yesterday
DockerShellAWSService MeshAzureCDNDNSDatadogGitGitHub ActionsKubernetesPythonRedisTerraform
Job Description
Job Title: Senior AWS Cloud Platform Engineer
Location: Naperville, IL
Work Model: Hybrid – Remote Monday & Friday | Onsite Tuesday through Thursday
Job Overview
We are seeking a highly skilled Senior AWS Cloud Platform Engineer with strong expertise in DevOps, Cloud Platform Engineering, Kubernetes, AWS, GitOps, CI/CD, and Infrastructure as Code.
The ideal candidate will be a hands-on platform engineer who can design, deploy, automate, and operate scalable cloud-native platforms. This role requires someone who continuously identifies opportunities to automate manual processes, improve platform reliability, reduce operational effort, and build internal engineering capabilities.
Key Responsibilities
Kubernetes & Container Platform Engineering
Strong hands-on experience with:
Location: Naperville, IL
Work Model: Hybrid – Remote Monday & Friday | Onsite Tuesday through Thursday
Job Overview
We are seeking a highly skilled Senior AWS Cloud Platform Engineer with strong expertise in DevOps, Cloud Platform Engineering, Kubernetes, AWS, GitOps, CI/CD, and Infrastructure as Code.
The ideal candidate will be a hands-on platform engineer who can design, deploy, automate, and operate scalable cloud-native platforms. This role requires someone who continuously identifies opportunities to automate manual processes, improve platform reliability, reduce operational effort, and build internal engineering capabilities.
Key Responsibilities
Kubernetes & Container Platform Engineering
- Deploy, manage, and troubleshoot workloads across AWS EKS and Azure AKS environments.
- Design, build, and maintain Kubernetes clusters throughout their lifecycle.
- Manage Kubernetes networking, ingress controllers, service meshes, pod communication, and traffic routing.
- Troubleshoot complex issues involving pods, nodes, networking, DNS, storage, and cluster scalability.
- Implement high-availability and disaster recovery strategies for Kubernetes platforms.
- Design and manage enterprise-scale GitOps workflows using Argo CD.
- Automate application deployments and environment promotion strategies.
- Build and maintain CI/CD pipelines using GitHub Actions.
- Define and enforce Git branching strategies and release management practices.
- Improve deployment reliability, rollback capabilities, and release governance.
Strong hands-on experience with:
- AWS EKS
- EC2
- Route 53
- IAM
- CloudFront/CDN
- ALB & NLB
- VPC and Cloud Networking
- Security Groups and NACLs
- AWS Secrets Manager
- ElastiCache (Redis)
- S3
- CloudWatch
- RDS
- Auto Scaling and High Availability Architectures
- Designing secure, scalable, and resilient AWS cloud architectures.
- Managing multi-account AWS environments.
- Implementing disaster recovery and business continuity solutions.
- Optimizing cloud cost, performance, security, and reliability.
- Build and maintain reusable Terraform modules.
- Provision and manage cloud infrastructure using Infrastructure as Code best practices.
- Maintain environment consistency and compliance through automation.
- Troubleshoot Terraform state issues, infrastructure drift, and large-scale deployments.
- Identify and eliminate manual operational processes through automation.
- Develop automation solutions for deployment, infrastructure, monitoring, and incident response.
- Build self-service capabilities and internal developer tools.
- Improve engineering productivity and reduce operational and tooling costs.
- Design monitoring and alerting solutions using Datadog or similar platforms.
- Develop automated alert correlation and incident reduction mechanisms.
- Create dashboards, SLOs, and observability standards.
- Drive proactive monitoring practices across production environments.
- Design highly resilient and fault-tolerant cloud architectures.
- Lead recovery planning for major production incidents and outages.
- Develop recovery strategies for region failures, Kubernetes cluster failures, account compromises, and infrastructure loss.
- Implement recovery using Infrastructure as Code, backups, replication, and automation.
- 4+ years of experience in DevOps, SRE, Platform Engineering, or Cloud Engineering.
- Expert-level Kubernetes administration experience.
- Strong hands-on experience with AWS EKS and/or Azure AKS.
- Deep understanding of Argo CD and GitOps principles.
- Strong experience with GitHub Actions and CI/CD automation.
- Advanced knowledge of Git branching and release management.
- Strong AWS architecture and cloud operations experience.
- Expert-level Terraform and Infrastructure as Code experience.
- Strong Docker and containerization experience.
- Linux system administration experience.
- Strong understanding of cloud networking and security.
- Azure cloud experience.
- Service Mesh experience.
- Multi-cloud deployment experience.
- Datadog or similar observability platform experience.
- Python, Go, Shell scripting, or other automation development experience.
- Strong Platform Engineering experience.
Similar jobs
- WS
DevOps Cloud Engineer with Security Clearance
White Sky Technologies
Annapolis Junction, MD🇺🇸$140k - $235k/yrHybrid5 weeks agoDockerRubyAWS+8Technology - WS
Senior DevOps Cloud Engineer with Security Clearance
White Sky Technologies
Annapolis Junction, MD🇺🇸$140k - $235k/yrHybrid6 weeks agoDockerRubyAWS+8Technology - EI
SRE PRODUCTION SUPPORT
NewEcho IT Solutions, Inc.
Alpharetta, GA🇺🇸On-siteYesterdayManufacturing - AS
AWS Lead DevOps Engineer
Apex Systems
Plano, TX🇺🇸On-site1 week agoMicroservicesAWSSnowflake+1Technology - RA
Web Application / DevOps Architect Onsite in San Jose, CA
NewRapidIT, Inc
San Jose, CA🇺🇸HybridYesterdayDockerFastAPIMemcached+16Technology - IC
Site Reliability Engineer (SRE)
Infinite Computer Solutions (ICS)
Frisco, TX🇺🇸Hybrid5 weeks agoDockerShellSplunk+11Technology