Haystack
← Back to Jobs
Technology
ES

Site Reliability Engineer

ElevaIT SolutionsCharlotte, NC🇺🇸United StatesPosted 27 Aug 2026

Why This Role Stands Out

Thrive in this hybrid role as a Site Reliability Engineer at a global leader in materials science, leveraging your expertise in Kubernetes and Linux to drive innovation in advanced computing infrastructure. You'll have the opportunity to enhance critical platforms, collaborate with cutting-edge research teams, and contribute to a company with a rich history of success. Embrace this chance to grow your skills and make a significant impact on a renowned organization.

Quick Overview

Seniority
Mid Senior
Work mode
Hybrid
Location
Charlotte, NC, United States
Posted
2 weeks ago
AWSScrumAgileArgoCDAzureGoogle CloudKanbanKubernetes

Job Description

Job Title: Site Reliability Engineer

Location: Remote (US, EST hours required)

Interview Process: 2x Interviews

Client Overview

This is a global materials science leader with a 170 plus year history, operating dozens of manufacturing and R&D sites worldwide. This role sits within a research and development group supporting advanced computing infrastructure behind ongoing materials science innovation.

Top 3 Skills

  • Kubernetes cluster operations and management, including provisioning, upgrades, and troubleshooting across on premises and cloud environments
  • Rancher for Kubernetes platform management
  • Linux systems administration, including performance tuning, scripting, and networking

What You'll Do

  • Maintain and enhance Kubernetes platforms across on premises and cloud environments
  • Support provisioning, upgrades, troubleshooting, and lifecycle management of Kubernetes clusters managed through Rancher
  • Provide deep Linux systems administration support, including performance tuning, troubleshooting, and automation
  • Develop and maintain infrastructure as code solutions to standardize and automate platform deployment
  • Support and improve GitOps workflows using ArgoCD to manage cluster and application configuration
  • Collaborate with developers, scientists, and infrastructure teams to deliver reliable platform services
  • Identify opportunities to improve platform resilience, observability, security, and maintainability

What We Need From You

  • 5 plus years of professional experience in site reliability engineering, platform engineering, DevOps, or systems engineering
  • Hands on experience operating Kubernetes platforms in production environments, both on premises and cloud based
  • Experience with Rancher for Kubernetes cluster management
  • Strong Linux systems administration skills, including troubleshooting, scripting, and system performance analysis
  • Experience implementing infrastructure as code solutions for platform provisioning and lifecycle management

Preferred/Bonus

  • Bachelor's degree in Computer Science, Software Engineering, Information Technology, or related field
  • Experience with Cluster API (CAPI)
  • Experience with hybrid infrastructure spanning on premises and public cloud (AWS, Azure, Google Cloud Platform)
  • Familiarity with Kubernetes observability, logging, monitoring, and alerting tooling
  • Experience supporting scientific research, high performance computing, or computational science environments
  • Experience with Agile teams (Scrum, Kanban)

Similar jobs