Quick Overview
Salary
$85 - $95/hr
Seniority
Leader
Work mode
Hybrid
Location
Pittsburgh, PA, United States
Posted
15 hours ago
AWSAzureGoogle CloudKubernetesTerraform
Job Description
Genesis10 is currently seeking a AVP / Head of Enterprise Site Reliability Engineering (SRE) with our consumer finance lender firm client in their Pittsburgh, PA location. This is a Right to hire position.
Summary:
An experienced Site Reliability Engineering leader is needed to establish, lead, and scale a formal enterprise SRE capability. This is a contract-to-hire opportunity intended to convert to a full-time AVP-level position based on performance, organizational approval, and mutually agreed-upon terms.
This leader will define the enterprise SRE strategy, operating model, governance framework, roadmap, and success measures while building and managing a centralized team focused on reliability, resilience, operational maturity, and engineering velocity across on-premises, hybrid, and cloud environments.
The role will drive the adoption of Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, observability standards, automation frameworks, incident management practices, and production-readiness expectations. The successful candidate will establish reliability as a measurable business and engineering outcome by creating repeatable patterns, paved-road standards, and operating practices that reduce toil and improve service performance.
This is a strategic, hands-on leadership position requiring strong people leadership, executive communication, program development, and cross-functional influence. The leader will partner with senior stakeholders across Technology, Product, Architecture, Security, Risk, Compliance, and Operations to make reliability a core product and platform capability.
Responsibilities:
Requirements:
Preferred Qualifications:
Ideal Candidate Profile:
Pay rate range: $85.00 - $95.00 hourly
If you have the described qualifications and are interested in this exciting opportunity, please apply!
Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals.
For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10:
For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website.
Genesis10 is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
Summary:
An experienced Site Reliability Engineering leader is needed to establish, lead, and scale a formal enterprise SRE capability. This is a contract-to-hire opportunity intended to convert to a full-time AVP-level position based on performance, organizational approval, and mutually agreed-upon terms.
This leader will define the enterprise SRE strategy, operating model, governance framework, roadmap, and success measures while building and managing a centralized team focused on reliability, resilience, operational maturity, and engineering velocity across on-premises, hybrid, and cloud environments.
The role will drive the adoption of Service Level Indicators (SLIs), Service Level Objectives (SLOs), error budgets, observability standards, automation frameworks, incident management practices, and production-readiness expectations. The successful candidate will establish reliability as a measurable business and engineering outcome by creating repeatable patterns, paved-road standards, and operating practices that reduce toil and improve service performance.
This is a strategic, hands-on leadership position requiring strong people leadership, executive communication, program development, and cross-functional influence. The leader will partner with senior stakeholders across Technology, Product, Architecture, Security, Risk, Compliance, and Operations to make reliability a core product and platform capability.
Responsibilities:
- Establish the enterprise SRE strategy, vision, roadmap, operating model, and governance framework in alignment with business priorities, technology modernization, cloud adoption, and operational resilience goals.
- Build, manage, and develop a centralized SRE team, including defining roles, creating staffing plans, establishing performance expectations, coaching team members, supporting career development, and planning for succession.
- Define the SRE engagement model, including service tiers, onboarding criteria, production-readiness standards, intake and prioritization processes, and partnership expectations for product, platform, infrastructure, and application teams.
- Define and operationalize enterprise reliability measures, including SLIs, SLOs, error budgets, availability targets, toil-reduction goals, incident metrics, change failure rate, mean time to detect, and mean time to restore.
- Establish error-budget policies and executive-level decision frameworks that guide tradeoffs among delivery velocity, operational risk, and service stability.
- Develop and govern enterprise reliability standards and reusable assets, including golden-signal frameworks, observability patterns, alerting standards, runbook templates, SLO dashboards, incident playbooks, and production-readiness checklists.
- Advance incident management maturity through effective high-severity response, executive communications, blameless post-incident reviews, root-cause analysis, corrective-action tracking, and measurable reliability improvements.
- Oversee operational readiness, capacity planning, disaster-recovery validation, resilience testing, service-health reviews, and reliability risk management for supported platforms and critical business services.
- Drive automation and observability strategies that reduce operational toil, improve visibility, accelerate recovery, and enable scalable support models across hybrid and cloud platforms.
- Identify, sponsor, and govern AI-enabled reliability capabilities, including alert-noise reduction, incident summarization, event correlation, runbook assistance, predictive operations, and approved automated remediation.
- Ensure AI-enabled operational capabilities are implemented with appropriate privacy, security, risk, compliance, auditability, and human-oversight controls.
- Partner with Security, Risk, Compliance, Audit, Architecture, Product, Development, Infrastructure, Cloud Operations, and Technology Operations leaders to embed reliability requirements throughout the software development lifecycle and production support model.
- Communicate reliability posture, risks, progress, business impact, and investment requirements to executive stakeholders using clear metrics and actionable recommendations.
- Promote shared ownership, engineering excellence, accountability, continuous improvement, and blameless learning across technology teams.
- Remain current on SRE, observability, platform engineering, AIOps, resilience engineering, and cloud reliability practices, applying relevant approaches to improve business and operational outcomes.
- Provide leadership during major incidents and critical operational events, including escalation management, cross-functional coordination, stakeholder communications, and executive updates.
Requirements:
- Progressive leadership experience in Site Reliability Engineering, infrastructure engineering, platform engineering, cloud operations, production engineering, technology operations, or a closely related discipline.
- Demonstrated success establishing a new SRE capability or significantly maturing an existing practice, including strategy, operating model, governance, service engagement, SLO management, production readiness, incident management, and roadmap execution.
- Proven experience hiring, managing, coaching, and developing engineers or other technical professionals.
- Strong executive presence and communication skills, with the ability to translate complex reliability, risk, and operational issues into clear business impacts, priorities, and decisions.
- Deep understanding of SRE practices, including SLIs, SLOs, error budgets, observability, incident command, post-incident reviews, toil reduction, automation, capacity planning, resilience testing, and production readiness.
- Experience leading high-severity incident response, coordinating cross-functional teams, communicating with senior stakeholders, and ensuring corrective actions are completed and evaluated for effectiveness.
- Strong knowledge of cloud and hybrid infrastructure, infrastructure as code, CI/CD, containerization, orchestration, monitoring, logging, distributed tracing, and modern production operations tooling.
- Ability to establish measurable programs, manage competing priorities, balance reliability with delivery velocity, and align technical investments with business criticality and risk reduction.
- Experience partnering with Security, Risk, Compliance, Audit, Architecture, Product, Development, Infrastructure, and Operations teams to establish secure, auditable, and sustainable reliability practices.
- Working knowledge of AIOps, predictive operations, AI-assisted incident response, observability automation, or related capabilities, including responsible-use and governance considerations.
- Only candidates available and ready to work directly as Genesis10 employees will be considered for this position.
- Education and Experience
- Bachelor's degree in Computer Science, Engineering, Information Technology, or a related discipline, or equivalent professional experience.
- Master's degree or comparable technology leadership experience preferred.
- At least 10 years of progressive technology experience, including five or more years in SRE, infrastructure engineering, platform engineering, cloud operations, production engineering, or technology operations leadership.
- At least three years of people-management experience leading engineers or technical teams, preferably within reliability, infrastructure, platform, cloud, or technology operations functions.
Preferred Qualifications:
- Experience establishing an enterprise SRE function within a large, complex, regulated, or highly distributed technology environment.
- Experience supporting business-critical services across on-premises, hybrid, and public-cloud platforms.
- Experience developing executive-level reliability reporting, service-health reviews, and investment recommendations.
- Relevant certifications in AWS, Microsoft Azure, Google Cloud, Terraform, Kubernetes, ITIL, DevOps, reliability engineering, automation, or technology leadership.
Ideal Candidate Profile:
- The ideal candidate combines enterprise strategy, technical credibility, people leadership, and operational judgment. This individual can build an SRE function from the ground up, establish practical governance without creating unnecessary friction, and influence senior leaders across multiple technology and business disciplines.
- The successful candidate will be comfortable operating at both strategic and execution levels-defining the long-term reliability model while helping teams address immediate operational risks, improve incident response, implement measurable SLOs, and create repeatable engineering standards.
Pay rate range: $85.00 - $95.00 hourly
If you have the described qualifications and are interested in this exciting opportunity, please apply!
Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals.
For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10:
- Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years.
- The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years.
- Access to an experienced, caring recruiting team (more than 7 years of experience, on average.)
- Behavioral Health Platform
- Medical, Dental, Vision
- Health Savings Account
- Voluntary Hospital Indemnity (Critical Illness & Accident)
- Voluntary Term Life Insurance
- 401K
- Sick Pay (for applicable states/municipalities)
- Commuter Benefits (Dallas, NYC, SF)
- Remote opportunities available
For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website.
Genesis10 is an Equal Opportunity Employer. Candidates will receive consideration without regard to their race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.
Similar jobs
- TC
Dynatrace SRE
NewTEKsystems c/o Allegis Group
Plano, TX🇺🇸$60 - $75/hrHybrid15 hours agoJiraTechnology - FB
Senior DevSecOps Engineer (Remote NC, TX, AZ, GA)
First-Citizens Bank & Trust Company
Austin, TX🇺🇸Remote4 days agoConfluenceJiraEngineering - FB
Senior DevSecOps Engineer (Remote NC, TX, AZ, GA)
First-Citizens Bank & Trust Company
Raleigh, NC🇺🇸Remote5 weeks agoConfluenceJiraEngineering - TC
Infrastructure Engineer
NewTEKsystems c/o Allegis Group
Decatur, IL🇺🇸$35/hrOn-site15 hours agoFiberTechnology - FB
Senior DevSecOps Engineer (Remote NC, TX, AZ, GA)
First-Citizens Bank & Trust Company
AZ🇺🇸Remote3 days agoConfluenceJiraEngineering - ID
Site Reliability Engineer
NewIDR, Inc.
Austin, TX🇺🇸HybridYesterdayActive DirectoryTechnology