Haystack
← Back to Jobs
Other

Application Management Services (AMS) Lead

Info Way SolutionsIrvine, CA🇺🇸United StatesPosted 28 Jul 2026

Why This Role Stands Out

You'll drive operational excellence and continuous improvement for mission-critical enterprise applications in a hybrid environment, offering excellent career growth. This role is perfect for a proactive leader with a strong background in incident management and observability, especially within the automotive domain, so seize this opportunity to make a significant impact.

Quick Overview

Work Type
Hybrid
Level
Mid Senior

Job Description

Job Summary

We are seeking an experienced Application Management Services (AMS) Lead to oversee production support, application monitoring, incident management, and operational excellence for mission-critical enterprise applications. The ideal candidate will have strong experience leading incident response, implementing observability solutions, and driving continuous service improvements in high-availability environments.

Experience in the Connected Car, Telematics, or Automotive domain is highly preferred.

Key Responsibilities

  • Lead end-to-end application monitoring and production support activities.
  • Own the complete incident management lifecycle including:
    • Detection
    • Triage
    • Resolution
    • Root Cause Analysis (RCA)
  • Lead critical incident (P1/P2) war rooms and coordinate with cross-functional teams.
  • Develop and maintain monitoring dashboards, alerts, and observability metrics.
  • Drive proactive monitoring initiatives to minimize outages.
  • Ensure SLA compliance and improve operational processes.
  • Create post-incident reports and preventive action plans.
  • Support on-call operations using escalation and notification tools.
  • Collaborate with engineering, infrastructure, DevOps, and product teams to improve system reliability.

Required Skills

Incident & Application Management
  • Application Management Services (AMS)
  • Production Support
  • Incident Management
  • Site Reliability Engineering (SRE)
  • Root Cause Analysis (RCA)
  • High Severity Incident Handling
  • Service Operations
Monitoring & Observability
  • Dynatrace
  • Grafana
  • ELK Stack (Elasticsearch, Logstash, Kibana)
  • OpenSearch
  • Monitoring Dashboards
  • Alerting & Metrics
ITSM & Collaboration
  • JIRA
  • Confluence
  • StatusPage
  • xMatters

Skills

ELK
Logstash
Compliance
Confluence
Grafana
Jira
Kibana
Root Cause Analysis
Telematics
Triage

Similar jobs