Why This Role Stands Out
This Principal SRE Architect role offers a unique opportunity to shape enterprise-wide operational resilience and cloud governance strategies, leading the integration of cutting-edge AI/ML workloads. You will thrive here if you are a visionary architect passionate about building robust, scalable, and secure systems, with the flexibility of a remote work environment. Apply now to make a significant impact on a global scale.
Quick Overview
Job Description
Job Title: Senior Principal SRE Architect
Location: Remote
Role Summary
As a Senior Principal SRE Architect, you will serve as the chief enterprise strategist for operational resilience, cloud governance, and modern infrastructure. You will define global SRE standards, drive chaos engineering and fault-tolerant architectural designs, and lead the operational integration of AI/ML workloads across multi-cloud environments.
Key Responsibilities
Architectural Leadership: Drive cross-domain SRE strategy, bridging enterprise front-end, back-end, API, database, and network architectures.
AI & Machine Learning Infrastructure: Architect scalable, cost-effective, and secure pipelines for deep learning, AI-native cloud services, and heavy compute workloads.
Chaos Engineering & Resilience: Establish proactive production-readiness criteria, automated failovers, chaos experiments, and risk mitigation strategies.
Enterprise CI/CD & DevOps Strategy: Define best practices for Infrastructure as Code (IaC), GitOps, and fully automated deployment workflows at scale.
Enterprise Security & Networking: Direct advanced network architecture, routing, firewall governance, zero-trust patterns, and cryptographic standards (ciphers, cert automation).
Core Technical Skills
Full-Stack Architecture: End-to-end enterprise systems design across APIs, Databases, Microservices, and Network Layers
Cloud & Containerization: AWS, Azure, Google Cloud Platform (including native AI services), Kubernetes, Container Ecosystems
AI/ML & Pipeline Architecture: TensorFlow, PyTorch, Deep Learning Frameworks, Scalable ML Operations (MLOps) Design
Operational Readiness & Resilience: Chaos Engineering, Production Readiness Reviews (PRRs), Enterprise Risk Mitigation
Observability & Telemetry: Enterprise-wide Logging, Distributed Tracing, Alerting Standards, Executive Dashboards
DevOps & CI/CD: Advanced CI/CD Strategy, Configuration as Code, Deployment Automation
Advanced Networking: TCP/IP, Wireshark, Enterprise Routing Protocols, Firewalls, F5/Load Balancers, Proxies, Automation Scripts
Security & Certificate Management: Venafi, ECMS, CerTIS, OpenSSL, Cipher Suite Optimization
IT Governance: ITSM Frameworks, Compliance, High Availability (HA), Disaster Recovery (DR) Strategy
Similar jobs
- SF
ServiceNow DevOps Lead Developer
NewSmart Folks Inc.
Chicago, IL🇺🇸Hybrid22 hours agoSOAPOAuthSonarQube+4Technology - CT
Director, Enterprise Architecture (DevOps)
NewCollinwood Technology Partners
Dallas, TX🇺🇸Hybrid22 hours agoMicroservicesAzureGoogle Cloud+1Technology - IA
Lead DevOps Engineer
NewIO Associates
New York, NY🇺🇸Hybrid22 hours agoAWSELKAzure+6Technology - SI
Google Cloud Platform DevOps Architect
NewSiri Infosolutions Inc
United States🇺🇸Remote22 hours agoFiberSQLAWS+6Technology - PN
Principal Site Reliability Engineer
PaloAlto Networks
CA🇺🇸$151.6k - $245.3k/yrOn-site1 month agoDockerMySQLShell+14Technology - AA
DEVOPS ENGINEER
AaraTechnologies Inc
Georgiana, AL🇺🇸Hybrid1 month agoDockerShellAWS+5Technology