Why This Role Stands Out
This hybrid Kubernetes Engineer role at Business Intelli Solutions Inc. offers significant opportunities for technical growth and impact by designing and managing robust EKS clusters. You'll thrive here if you have a passion for Kubernetes architecture, Infrastructure as Code, and optimizing cloud-native environments, so be sure to apply!
Quick Overview
Seniority
Mid Senior
Work mode
Hybrid
Location
TX, United States
Posted
2 weeks ago
AWSService MeshBashDNSGrafanaHelmKubernetesPrometheusPythonTerraform
Job Description
Kubernetes Engineer role
Key Responsibilities
- Design, provision, and manage production-grade Amazon EKS clusters across development, test, and production environments using Infrastructure as Code.
- Architect multi-cluster, multi-AZ Kubernetes environments with high availability, fault tolerance, and workload isolation.
- Manage full cluster lifecycle - provisioning, version upgrades, node group management, add-on lifecycle, and decommissioning.
- Design and implement Kubernetes networking - VPC CNI configuration, ingress controllers (AWS Load Balancer Controller), service mesh, and DNS.
- Implement and manage cluster autoscaling strategies using Karpenter and/or Cluster Autoscaler, including Spot/On-Demand node pool design and cost optimization.
- Manage role-based access control (RBAC) - create and maintain Roles, ClusterRoles, RoleBindings, and ClusterRoleBindings across namespaces and teams.
- Implement IRSA (IAM Roles for Service Accounts) to assign AWS permissions to workloads without node-level credentials.
- Build and maintain Infrastructure as Code (Terraform) for all cluster and supporting AWS infrastructure - VPC, IAM, ECR, and EKS.
- Define and enforce resource governance - ResourceQuotas, LimitRanges, PodDisruptionBudgets, and PriorityClasses across namespaces and teams.
- Design and manage persistent storage - EBS CSI and EFS CSI drivers, StorageClass configuration, and backup strategies.
- Build and maintain Dynatrace dashboards to provide operational visibility into cluster health, workload performance, and infrastructure metrics.
- Configure Dynatrace monitors and alerts for proactive issue detection across Kubernetes infrastructure.
- Collaborate with DevOps, SRE, and Application teams to onboard workloads, define cluster standards, and resolve platform-level incidents.
- Perform root cause analysis and resolve production incidents related to cluster infrastructure, scheduling, networking, and storage.
Required Qualifications
- Bachelor's degree in Computer Science, Information Technology, Engineering, or related field.
- 7+ years of overall IT Infrastructure, Cloud, or Platform Engineering experience.
- 4+ years of hands-on Kubernetes cluster administration in production environments.
- Strong experience designing, building, and operating Amazon EKS clusters end-to-end.
- Deep knowledge of Kubernetes networking - VPC CNI, ingress, and DNS.
- Hands-on experience with Karpenter or Cluster Autoscaler for dynamic node provisioning.
- Strong Terraform experience for EKS and AWS infrastructure provisioning.
- Experience with Kubernetes RBAC and IRSA for role and permission management.
- Hands-on experience building Dynatrace dashboards and configuring monitors and alerts.
- Experience with CI/CD systems and GitOps tooling for workload delivery.
- Linux administration and AWS core services - VPC, IAM, EC2, S3, ECR.
Preferred Skills
- AWS certifications - Solutions Architect, DevOps Engineer, or equivalent.
- CKA (Certified Kubernetes Administrator) certification.
- Scripting proficiency in Python and/or Bash for operational automation.
- Experience with Helm or Kustomize for workload delivery.
- Additional observability tooling - Prometheus, Grafana.
- Multi-cloud or hybrid cluster experience (AKS, GKE alongside EKS).
Similar jobs
- JM
Infrastructure Engineer III - Site Reliability
NewJ.P. Morgan
Plano, Texas🇺🇸On-site31 minutes agoGCPAWSSplunk+6Technology - OR
Senior Manager, Site Reliability Engineering
NewOracle
Reston, Virginia🇺🇸$121.5k - $264.1k/yrHybrid39 minutes agoOracleTechnology - BA
SRE Virtual Desktop Operations Engineer - AVP
NewBarclays
New York City, New York🇺🇸Hybrid39 minutes agoSplunkActive DirectoryAnsible+7Technology - GE
Site Reliability Engineer
NewGenesis10
Plano, TX🇺🇸$64 - $72/hrHybridYesterdayAgileAnsibleGit+2Technology - AS
DevOps Engineer - Middleware/Messaging Focus
NewApex Systems
Round Rock, TX🇺🇸HybridYesterdayDockerShellFlink+11Technology - GE
Site Reliability Engineer
NewGenesis10
Chandler, AZ🇺🇸$60 - $68/hrHybridYesterdaySOAPSQLShell+14Technology