Haystack
← Back to Jobs
Technology
SD

Observability Operations Engineer

Stanley David and AssociatesPhoenix, AZ🇺🇸United StatesPosted Aug 31, 2026

Why This Role Stands Out

Advance your career as an Observability Operations Engineer at Stanley David and Associates, where you'll gain valuable experience with leading observability tools and contribute to critical technology operations in a hybrid work environment. This role is perfect for a mid-senior level professional eager to develop their skills within a reputable company and a supportive team. Apply today to explore this exciting opportunity!

Quick Overview

Seniority
Mid Senior
Work mode
Hybrid
Location
Phoenix, AZ, United States
Posted
4 weeks ago
ShellSplunkGrafanaKubernetesPhoenixPrometheusPythonREST

Job Description

Role         ::         Observability Operations Engineer

Location ::         Phoenix, AZ 

Type        ::          Fulltime

 

Job Description

 

Must Have Technical/Functional Skills

•                          Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch. 

•                          Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning. 

•                          Strong knowledge of Linux, Kubernetes, and cloud environments. 

•                          Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies. 

•                          Automation experience using Python/Shell scripting and REST APIs. 

•                          Experience supporting enterprise-scale production environments, troubleshooting, and RCA.

 

Roles & Responsibilities

•                          Administer and optimize enterprise Dynatrace, Splunk, and OpenSearch/Elasticsearch platforms.

•                          Maintain platform availability, scalability, performance, security, and reliability.

•                          Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.

•                          Troubleshoot production issues and perform root cause analysis using observability tools.

•                          Support Linux, Kubernetes, container, and cloud-based environments.

•                          Automate operational activities and drive self-healing and AI-assisted operations.

•                          Manage upgrades, patching, capacity planning, backups, and operational governance.

•                          Collaborate with SRE, DevOps, Platform, Infrastructure, and Application teams.

 

Generic Managerial Skills, If any

 Good Communication and assertiveness, Team Player

Similar jobs