Haystack
← Back to Jobs
Remote
Technology

Site Reliability Engineer – Kubernetes Platform (FedRAMP High / IL5)

SDH SystemsUnited States🇺🇸United StatesPosted 23 Jul 2026

Why This Role Stands Out

Advance your career in a fully remote Site Reliability Engineer role, focusing on cutting-edge Kubernetes platforms within regulated environments, offering significant growth and impact. You'll thrive here if you're passionate about building scalable, reliable infrastructure and collaborating with a talented team to ensure platform excellence. Don't miss this opportunity to contribute to critical systems and enhance your expertise!

Quick Overview

Work Type
Remote
Level
Mid Senior

Job Description

Job Title: Site Reliability Engineer (SRE) – Kubernetes Platform (FedRAMP High / IL5)

Location: 100% Remote

Interview Process: Online

Work Schedule: 100% Remote

Job Description:

Level: L1

Location: Remote, prefer PST hours

About the Team 

The SRE Platform Engineering team builds and operates the infrastructure that powers our cloud. We focus on delivering reliable, scalable, and simple platforms that enable product teams to move quickly while meeting the requirements of regulated environments such as FedRAMP High and DoD IL5. 

 

About the Role 

We’re looking for a Site Reliability Engineer to support the development and operation of our Kubernetes-based platform in regulated environments. In this role, you will work closely with senior engineers and technical leaders to improve reliability, scalability, and compliance across the platform. 

This is a hands-on engineering role where you’ll contribute to key systems, build infrastructure to support production operations, and help implement solutions that improve the overall health and performance of the platform. 

 

What You Will Do 

  • Contribute to the design, implementation, and operation of Kubernetes platforms in FedRAMP High / IL5 environments 

  • Support day-to-day reliability and performance of platform services, including monitoring and alerting 

  • Cisco Confidential 



  • Implement automation and tooling to improve operational efficiency and reduce manual effort 

  • Work with senior engineers to define and track SLIs, SLOs, and error budgets 

  • Assist in maintaining compliance and security requirements, including support for audits and continuous monitoring 

  • Contribute to infrastructure as code and CI/CD pipeline improvements 

  • Collaborate with cross-functional teams (Security, Platform, Application teams) to resolve issues and deliver platform capabilities 

  • Participate in on-call rotations supporting customer requests and paging alerts 

 

What You Bring 

  • 4–6 years of experience in SRE, DevOps, or platform engineering roles 

  • Experience with Kubernetes in production environments 

  • Familiarity with cloud platforms (AWS, Azure, or similar; GovCloud experience a plus) 

  • Solid understanding of Linux systems, networking, and containerization 

  • Experience with Infrastructure as Code (e.g., Terraform) 

  • Proficiency in scripting or programming (e.g., Python, Go) 

  • Exposure to observability tools (Prometheus, Grafana, logging systems) 

 

Nice to Have 

  • Experience working in FedRAMP High or DoD IL5 environments 

  • Exposure to CI/CD systems and deployment automation (e.g., ArgoCD) 

  • Familiarity with container security practices and tools 

  • Experience supporting regulated or audited systems 

 

How You Work 

  • You take ownership of your work while seeking guidance when needed 

  • You collaborate effectively with more senior engineers and technical leaders 

  • You focus on building reliable, maintainable solutions 

  • You are comfortable working in structured, compliance-driven environments 

  • You are proactive about learning and improving systems 

Skills

AWS
ArgoCD
Azure
Grafana
Kubernetes
Prometheus
Python
Terraform

Similar jobs