Quick Overview
Seniority
Leader
Work mode
Hybrid
Location
Chicago, IL, United States
Posted
2 days ago
AWSDatadogC++GrafanaJavaKubernetesPythonTerraform
Job Description
Principal Site Reliability Engineer
An established fintech institution is seeking a Principal Site Reliability Engineer (SRE) to apply software engineering and systems engineering best practices to ensure the stability, scalability, and overall health of critical production services. This full-time position is hybrid, with 3 days on-site in Downtown Chicago and 2 days remote.
This is a highly influential Principal-level role with significant impact on engineering and reliability strategy. This role collaborates with engineering and technology teams to build reliable, scalable, and observable systems throughout their lifecycle. The Principal SRE leads technical initiatives across the enterprise by leveraging data-driven engineering, automation, and strong architectural expertise to solve complex challenges and align teams around scalable solutions.
Required Skills & Experience
Tech Breakdown
This role is bonus eligible.
You will receive the following benefits:
An established fintech institution is seeking a Principal Site Reliability Engineer (SRE) to apply software engineering and systems engineering best practices to ensure the stability, scalability, and overall health of critical production services. This full-time position is hybrid, with 3 days on-site in Downtown Chicago and 2 days remote.
This is a highly influential Principal-level role with significant impact on engineering and reliability strategy. This role collaborates with engineering and technology teams to build reliable, scalable, and observable systems throughout their lifecycle. The Principal SRE leads technical initiatives across the enterprise by leveraging data-driven engineering, automation, and strong architectural expertise to solve complex challenges and align teams around scalable solutions.
Required Skills & Experience
- 15+ years of professional experience in Software Engineering, Site Reliability Engineering, Systems Engineering, Cloud/Platform Engineering, DevOps, or Infrastructure Engineering.
- Strong leadership capabilities including the ability to mentor junior engineers and the capacity to drive architecture decisions and influence technical direction.
- The ability to operates with significant autonomy across the organization's most consequential reliability challenges.
- Proven track record of using objective analysis and technical expertise to validate assumptions, assess risks, and guide engineering decisions
- AWS/Cloud experience
- C++ experience
- Bachelor's degree in Computer Science, Software Engineering, Computer Engineering, or Information Systems
- Experience in the Security, Payments, or Financial Services industries
Tech Breakdown
- Languages/scripting: Java, Python
- Containers/clusters: Kubernetes
- IaC: Terraform
- Observability/monitoring: Datadog, Grafana
- OS: Linux
- Apply automation, DevOps, and software engineering best practices to improve how services are built, deployed, monitored, and maintained.
- Use data and analysis to identify reliability risks, validate solutions, and guide technical decisions.
- Define and enhance service reliability metrics, objectives, and performance standards.
- Improve system observability through monitoring, logging, tracing, alerting, and dashboards.
- Drive continuous improvements in CI/CD, Infrastructure as Code, automation, testing, incident response, and system reliability.
- Identify recurring production issues and implement long-term improvements in code, architecture, tooling, and operational processes.
This role is bonus eligible.
You will receive the following benefits:
- Competitive medical (PPO and HDHP), dental, and vision insurance, plus employer contributions to Health Savings Accounts (HSA) and Flexible Spending Accounts (FSA) for healthcare, commuting, and dependent care expenses.
- Dollar-for-dollar 401(k) match on the first 6% of employee contributions, available upon eligibility.
- Flexible Time Off (FTO) for salaried employees, generous PTO for hourly employees, 11 paid company holidays, and a paid volunteer day.
- Up to 12 weeks of paid parental leave to support growing families.
- Access to Maven's comprehensive family planning benefits, including fertility treatments, egg freezing, adoption, surrogacy, pregnancy and postpartum care, pediatric support, and return-to-work resources.
Similar jobs
- MR
DevSecOps Engineer
NewMotion Recruitment Partners, LLC
Boston, MA🇺🇸HybridYesterdayEngineering - AI
Lead Azure DevOps Engineer with Banking & Finance
NewAmerican IT Systems
United States🇺🇸Remote8 hours agoGCPAWSETL+3Technology - MR
Senior Cloud Platform Engineer
NewMotion Recruitment Partners, LLC
United States🇺🇸RemoteYesterdayShellAWSAzure+10Technology - RH
Site Reliability Engineer (SRE)
Robert Half
Maumee, OH🇺🇸Hybrid3 weeks agoAzureDatadogGitHub Actions+4Technology - JG
Infrastructure Automation Engineer
NewJudge Group, Inc.
Reston, VA🇺🇸HybridYesterdayAnsibleBashDatadog+4Technology - MR
Senior Platform Engineer/ High Growth
NewMotion Recruitment Partners, LLC
Toronto, ON🇺🇸HybridYesterdayDynamoDBRubyShell+5Technology