Haystack
← Back to Jobs
Technology
RS

Lead DevOps/Site Reliability Engineer -- FTE-- Onsite at FL

REDLEO SOFTWARE INC.Orlando, FL🇺🇸United StatesPosted 25 Aug 2026

Why This Role Stands Out

This Lead Site Reliability Engineer role offers a fantastic opportunity to shape the infrastructure behind cutting-edge AI systems at a reputable company. You'll thrive here if you have a strong background in SRE, Kubernetes, and Terraform, with a keen interest in AI. Apply now to leverage your expertise and drive innovation in this exciting on-site position.

Quick Overview

Seniority
Mid Senior
Work mode
On Site
Location
Orlando, FL, United States
Posted
2 weeks ago
DockerAWSAzureGoogle CloudHelmKubernetesTerraform

Job Description

Lead Site Reliability Engineer – Infrastructure & DevOps

Location: Onsite – Orlando, FL (Open to Glendale/Anaheim, Seattle, Orlando)
Category: DevOps / SRE

CLIENT NOTE:

Candidates should have some exposure to AI systems and infrastructure. They should have experience supporting the infrastructure behind AI systems or implementing visibility, observability, and monitoring to assess AI system health and performance.


MUST HAVE:

  • 7+ years SRE / Platform / DevOps experience
  • Expert Kubernetes, Terraform, Helm
  • Multi‑cloud production experience: AWS, Google Cloud Platform, Azure
  • Proven technical leadership in high‑availability environments
  • Strong CI/CD automation (Harness or equivalent)
  • Deep observability experience (monitoring, logging, alerting)
  • Experience supporting mission‑critical, highly available production systems

Required Skills

  • AWS, Google Cloud Platform cloud networking
  • Docker, Kubernetes
  • Terraform, Helm
  • CI/CD pipelines (Harness preferred)
  • Multi‑cloud operations
  • Infrastructure automation & reliability engineering

Role Overview

  • Lead SRE initiatives across multi‑cloud environments
  • Architect, automate, and optimize infrastructure for reliability and scalability
  • Build and maintain CI/CD pipelines and deployment automation
  • Enhance observability, monitoring, and alerting systems
  • Support production workloads, troubleshoot incidents, and ensure uptime
  • Collaborate with engineering, platform, and security teams
  • Drive best practices for infrastructure, DevOps, and cloud operations

Similar jobs