Quick Overview
Job Description
We are seeking a highly motivated Production Operations (ProdOps) Engineer to manage and support mission-critical production platforms. The ideal candidate will have strong expertise in GitLab CI/CD, Kubernetes, PostgreSQL, Database Deployments, Application Troubleshooting, Dynatrace Monitoring, and On-Call Production Support. This role requires a proactive individual who can ensure platform stability, automate operational processes, improve deployment efficiency, and rapidly resolve production incidents.
Key Responsibilities
Provide production/non-production support and participate in 24x7 on-call.
Monitor systems, troubleshoot incidents, and perform RCA.
Build and optimize GitLab CI/CD pipelines and deployment automation.
Manage Kubernetes, Helm, pods, services, and ingress.
Handle database deployments, migrations, rollbacks, and PostgreSQL administration.
Configure Dynatrace monitoring, APM, alerts, and observability.
Develop automation, runbooks, and processes to improve platform reliability and efficiency.
Position Requirements:
Technical Skills
GitLab CI/CD
Pipeline creation and optimization
GitLab Runners
Deployment automation
Kubernetes
Cluster operations and administration
Helm Charts
Pod, Service, Ingress troubleshooting
Database Deployments
Schema migrations
Release management
Rollback strategies
Monitoring & Observability
Dynatrace
Application Performance Monitoring (APM)
Log analysis and alerting
Troubleshooting
Application debugging
Infrastructure issue diagnosis
Performance analysis
Production incident management
Operational Skills
24x7 On-Call Support
Incident Management
Root Cause Analysis (RCA)
Problem Management
Change Management
Release Coordination
Preferred Qualifications
Bachelor’s degree in computer science, Information Technology, or related field.
5+ years of experience in Production Support, DevOps, SRE, or Platform Engineering roles.
Experience with Linux administration and shell scripting.
Familiarity with cloud platforms such as Azure, AWS, or Google Cloud Platform.
Experience with Infrastructure as Code (Terraform preferred).
Understanding of microservices architecture and containerization technologies.
Success Metrics
Production availability and up time.
Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
Deployment success rate.
Reduction in recurring incidents.
Platform reliability and performance improvements.
Operational automation and efficiency gains.
Similar jobs
- RT
Principal Production Test Engineer with Security Clearance
RTX
Huntsville, AL🇺🇸$107.5k - $204.5k/yrHybrid3 weeks agoManufacturing - KT
Production Supervisor 5
NewKforce Technology Staffing
Wentzville, MO🇺🇸On-siteYesterdayManufacturing - SF
Fire Protection Sales / Service Technician
NewSSI Fire & Safety Holdings, LLC.
United States🇺🇸$40k - $60k/yrOn-siteYesterdayBusiness DevelopmentMicrosoft OfficeManufacturing - SA
Automotive Service Technician / Internal Technician - Frontier Leasing and Sales
NewSonic Automotive
United States🇺🇸$25 - $40/hrOn-siteYesterdayManufacturing - BR
Mercury Marine: Human Resources Intern - Manufacturing Operations
NewBrunswick Corporation
United States🇺🇸$18 - $28/hrOn-site8 hours agoManufacturing - JA
Human Resources Generalist III (Manufacturing)
NewJabil
United States🇺🇸$79.4k - $142.9k/yrOn-site8 hours agoContinuous ImprovementManufacturing