Haystack
← Back to Jobs
Technology

Site Reliability Engineer(Observability Engineer )

Key2Source INCSan Francisco, CA🇺🇸United StatesPosted 6 Aug 2026

Quick Overview

Work Type
On Site
Level
Mid Senior

Job Description

Role: Site Reliability Engineer, Observability Engineer (AI Tools, AppDynamics & Splunk)

Location: San Francisco, CA (Onsite)

Job Summary

  • We are seeking an experienced Observability Engineer to design, implement, and manage enterprise monitoring and observability solutions. The ideal candidate will have hands-on experience with AI-driven observability tools, AppDynamics, and Splunk to ensure application performance, system reliability, and proactive incident management.

Responsibilities

  • Implement and maintain enterprise observability and monitoring solutions.
  • Configure and manage AppDynamics and Splunk dashboards, alerts, and reports.
  • Utilize AI-powered monitoring tools for anomaly detection and root cause analysis.
  • Monitor application, infrastructure, and cloud environments to ensure high availability.
  • Collaborate with DevOps, Infrastructure, and Application teams to troubleshoot production issues.
  • Create monitoring standards, documentation, and operational runbooks.
  • Required Skills
  • Experience with AppDynamics and Splunk.
  • Knowledge of AI-powered observability and monitoring platforms.
  • Strong understanding of application performance monitoring (APM).
  • Experience with cloud platforms (AWS, Azure, or Google Cloud Platform).
  • Familiarity with Linux, scripting (Python/Shell), and CI/CD environments.
  • Excellent troubleshooting and communication skills.

Skills

Shell
AWS
Splunk
Azure
Google Cloud
Python

Similar jobs