Haystack
← Back to Jobs
Manufacturing
IF

DevOps Engineers with Production Support || NJ/OH

IT First SourceJersey City, NJ🇺🇸United StatesPosted Sep 24, 2026

Why This Role Stands Out

This role offers significant impact by ensuring the reliability and performance of a critical enterprise browser platform, providing you with excellent opportunities for technical growth and problem-solving. If you thrive on engineering solutions to complex operational challenges and possess strong Linux administration and scripting skills, you'll excel here. Apply today to join a team focused on innovation and robust production support.

Quick Overview

Seniority
Mid Senior
Work mode
On Site
Location
Jersey City, NJ, United States
Posted
22 hours ago
Root Cause Analysis

Job Description

Secure Enterprise Browser DevOps Engineers with Production Support experience

Jersey City, NJ or Columbus, OH

Role Overview

As an Enterprise Browser DevOps Engineer, you will be responsible for the reliability, availability, and performance of the firm's secure enterprise browser platform. You will treat operations as a software problem: define service levels, engineer away toil, and keep the platform production-ready as it scale.

Reporting to the platform's lead engineer, you will share production ownership and on-call duties with SRE and incident response teams, removing single points of failure on a system people depend on to do their jobs. You will lead incident response, root cause analysis, harden the platform against mass-impact events, and build the automation, monitoring, and access controls that let a small team run a large environment safely.

Job Responsibilities

  • Engineer, test, and deploy access and identity controls group policy and SAML-based role
  • Manage safe change through browser version and extension lifecycle, including staged rollout, validation, and fast rollback.
  • Eliminate toil by building automation for deployment, configuration, monitoring, and health checks.
  • Define service-level objectives and deliver the observability, performance analytics, and capacity insight that keep the platform healthy at scale.
  • Production reliability and lead incident response during outages and mass-impact events, driving tier-4 troubleshooting and blameless post-incident review to permanent fixes.

Required Qualifications, Capabilities, and Skills

  • Operated business-critical production systems, owning reliability, availability, and on-call.
  • Software-driven operations: Linux administration and coding (Python, Bash) with infrastructure-as-code and configuration management.
  • Delivered endpoint or enterprise browser controls, including group policy, SAML/SSO, and extension governance.
  • Built CI/CD, observability, and automated health checks that reduce toil across the deployment lifecycle.
  • Led incident command and root-cause analysis, turning findings into lasting reliability improvements.

Thanks, IT First Source

Similar jobs