Senior DevOps Architect || Remote
Why This Role Stands Out
This remote Senior DevOps Architect role offers a unique opportunity to shape enterprise-wide, multi-cloud strategies and build innovative, AI-powered infrastructure from the ground up. You'll thrive here if you're a visionary leader passionate about automation, cost optimization, and mentoring engineering teams, driving significant impact within a reputable organization.
Quick Overview
Job Description
Architecture, Design & Innovation
- Enterprise Infrastructure & Multi-Cloud Strategy: Lead the end-to-end architecture of scalable, secure, and high-performing DevOps and Platform Engineering solutions across multi-cloud environments (AWS, Azure, and Google Cloud Platform), as well as specialized infrastructures like VMC on AWS (VMware Cloud) and GovCloud.
- Reference Architecture & Automation: Define and own the DevOps reference architecture guardrails. Design and implement cutting-edge automated frameworks, including advanced patterns like Bot-as-a-Service (BaaS) and AIOps platforms for predictive operations.
- AI-Powered Workloads: Formulate structural guardrails for infrastructure supporting complex AI/ML workloads (GPU compute, vector databases, model serving, and pipeline orchestration).
Toolchain, Frameworks & Cloud Cost Optimization
- CI/CD from Scratch: Oversee the strategy, design, and continuous improvement of enterprise-wide CI/CD pipelines and release frameworks from scratch, establishing standardized delivery rails across globally distributed cross-functional teams.
- Infrastructure as Code (IaC): Champion robust IaC methodologies using Terraform, Ansible, or Pulumi across multi-region ecosystems.
- Multi-Million Dollar Cost Governance: Own infrastructure monitoring and utilization strategies; proactively lead resource cost planning and optimization initiatives to drive significant efficiency gains and eliminate multi-million dollar cloud cost leaks.
Technical Leadership, Mentorship & Client Engagement
- Center of Excellence: Act as a core mentor and coach for engineering teams, including Engineering Managers and Tech Leaders, promoting a culture of innovation, research, and operational excellence.
- SaaS Operations & Upgrade Governance: Guide teams through the operational lifecycle of deploying the SaaS platform, managing complex patching, end-to-end zero-downtime upgrades, and release cadences across multiple hybrid or complex client environments.
- Escalation & Stakeholder Management: Effectively handle high-stakes client engagements and manage critical operational escalations, seamlessly translating complex technical resolutions into clear strategies for both engineering groups and executive leaders.
- Hiring & Culture: Drive technical and behavioral interviewing standards to evaluate top-tier talent and grow the technical function.
Monitoring, Reliability & Security
- Observability Solutions: Architect extensive distributed tracing, centralized logging, and alerting systems to manage platform health proactively.
- Regulated Compliance: Enforce strict SLIs, SLOs, and error budgets while embedding robust DevSecOps guardrails compliant with SOC 2, ISO 27001, and rigid Federal/GovCloud regulatory requirements.
What you'll have
Experience & Profile Requirement
- 15+ Years of Experience: A minimum of 15 years of progressive experience in DevOps, Platform Engineering, Build & Release, Infrastructure, or Site Reliability Engineering (SRE) roles.
- Senior Architectural Leadership: At least 5 7 years serving in a Principal Architect, Principal Engineer, or Technical Manager capacity influencing cross-functional organizational architectures.
- Deep Cloud & Hybrid Footprint: Demonstrated mastery in orchestrating global-scale SaaS applications across AWS, Azure, and Google Cloud Platform, alongside specialized experience handling VMC on AWS and GovCloud environments.
- Cost Governance Track Record: Proven background in designing resource monitoring solutions that have successfully delivered multi-million-dollar cloud cost optimization and reduction metrics.
- Release & Lifecycle Management Mastery: Comprehensive experience managing end-to-end upgrade processes, emergency patching, and intricate multi-tenant deployment structures for both cloud-native and highly isolated on-premises client environments.
Skills
Similar jobs
AWS Devops Engineer - W2
Intellisoft Technologies · Plano, United States
6 minutes agoCloud Data Platform Engineer with Snowflake & Redshiftplatform
Gemini Consulting Services · United States
6 minutes agoAssociate Platform Engineer
Kforce Technology Staffing · West Palm Beach, United States
6 minutes agoAI Platform Engineer
Ace Technologies, Inc. · San Francisco, United States
8 minutes agoAWS Software Developer (Cloud / DevOps) w/ CI Poly with Security Clearance
Strategic Business Systems · Chantilly, United States
11 minutes agoSalesforce Platform Engineer
TEKEngineersInc · Milwaukee, United States
49 minutes ago