Why This Role Stands Out
Thrive in this hybrid role as a Site Reliability Engineer at a global leader in materials science, leveraging your expertise in Kubernetes and Linux to drive innovation in advanced computing infrastructure. You'll have the opportunity to enhance critical platforms, collaborate with cutting-edge research teams, and contribute to a company with a rich history of success. Embrace this chance to grow your skills and make a significant impact on a renowned organization.
Quick Overview
Job Description
Job Title: Site Reliability Engineer
Location: Remote (US, EST hours required)
Interview Process: 2x Interviews
Client Overview
This is a global materials science leader with a 170 plus year history, operating dozens of manufacturing and R&D sites worldwide. This role sits within a research and development group supporting advanced computing infrastructure behind ongoing materials science innovation.
Top 3 Skills
- Kubernetes cluster operations and management, including provisioning, upgrades, and troubleshooting across on premises and cloud environments
- Rancher for Kubernetes platform management
- Linux systems administration, including performance tuning, scripting, and networking
What You'll Do
- Maintain and enhance Kubernetes platforms across on premises and cloud environments
- Support provisioning, upgrades, troubleshooting, and lifecycle management of Kubernetes clusters managed through Rancher
- Provide deep Linux systems administration support, including performance tuning, troubleshooting, and automation
- Develop and maintain infrastructure as code solutions to standardize and automate platform deployment
- Support and improve GitOps workflows using ArgoCD to manage cluster and application configuration
- Collaborate with developers, scientists, and infrastructure teams to deliver reliable platform services
- Identify opportunities to improve platform resilience, observability, security, and maintainability
What We Need From You
- 5 plus years of professional experience in site reliability engineering, platform engineering, DevOps, or systems engineering
- Hands on experience operating Kubernetes platforms in production environments, both on premises and cloud based
- Experience with Rancher for Kubernetes cluster management
- Strong Linux systems administration skills, including troubleshooting, scripting, and system performance analysis
- Experience implementing infrastructure as code solutions for platform provisioning and lifecycle management
Preferred/Bonus
- Bachelor's degree in Computer Science, Software Engineering, Information Technology, or related field
- Experience with Cluster API (CAPI)
- Experience with hybrid infrastructure spanning on premises and public cloud (AWS, Azure, Google Cloud Platform)
- Familiarity with Kubernetes observability, logging, monitoring, and alerting tooling
- Experience supporting scientific research, high performance computing, or computational science environments
- Experience with Agile teams (Scrum, Kanban)
Similar jobs
- ST
Lead DevOps Engineer
NewStefanini
Richmond, VA🇺🇸Remote23 hours agoJenkinsTechnology - BT
Cloud DevOps Engineer – Azure
NewBrillfy Technology
United States🇺🇸Hybrid23 hours agoDockerPackerShell+8Technology - BO
Build Reliability Engineer - Millennium Space Systems with Security Clearance
Boeing
El Segundo, CA🇺🇸$107.1k - $157.5k/yrHybrid4 weeks agoAssemblyTechnology - PS
Mid-Level DevOps Software Engineer, SE2 with Security Clearance
NewPower3 Solutions
Hanover, MD🇺🇸$206k - $239k/yrHybrid2 days agoMongoDBLogstashAgile+9Technology - AS
CI/CD Cloud Platform Engineer
NewApex Systems
Milwaukee, WI🇺🇸Hybrid23 hours agoAWSELKCloudFormation+8Technology - AS
Site Reliability Engineer
NewApex Systems
Greenwood Village, CO🇺🇸Hybrid23 hours agoNode.jsSQLAWS+14Technology