Haystack
← Back to Jobs
Technology
VC

Site Reliability Engineer III

VST Consulting, IncPennington, NJ🇺🇸United StatesPosted Sep 30, 2026

Quick Overview

Salary
$60 - $63/hr
Seniority
Mid Senior
Work mode
On Site
Location
Pennington, NJ, United States
Posted
19 hours ago
DjangoMySQLAWSAnsibleAzurePerlPythonREST

Job Description

Role : Site Reliability Engineer III Location: Pennington, NJ 08534
Work Type: Onsite / Local Candidates Preferred
Employment Type: W2
Visa: Any Visa is accepted
Pay Rate: $60 - 63/hr on W2

Job Description

We are seeking a highly skilled Site Reliability Engineer (SRE) III to work hands-on across the technology stack, improving platform and application reliability, observability, and operational efficiency across Global Markets.

The SRE will work closely with engineering and core technology teams to support platform modernization, cloud migration initiatives, production system hardening, monitoring improvements, and operational tooling. This position is critical to maintaining execution velocity, reducing operational risk, and ensuring platforms meet reliability and performance objectives.

Responsibilities

  • Design, implement, and maintain reliable, scalable, and highly available production systems.
  • Define, implement, and monitor Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
  • Perform performance engineering and analysis using Dynatrace and modern observability tools.
  • Develop automation and operational tooling using Python or Perl.
  • Develop and maintain applications and REST APIs using Python and Django.
  • Troubleshoot complex issues across Linux, applications, databases, infrastructure, and cloud environments.
  • Develop and maintain infrastructure automation using Ansible and Infrastructure-as-Code frameworks.
  • Build, maintain, and improve CI/CD pipelines and DevOps automation.
  • Support cloud migration and modernization initiatives across AWS and/or Azure environments.
  • Implement and maintain observability solutions using Dynatrace, OpenTelemetry, and related monitoring technologies.
  • Analyze production performance, identify reliability risks, and implement corrective actions.
  • Improve system availability, scalability, resilience, and operational efficiency.
  • Collaborate with development, infrastructure, cloud, and platform teams to resolve production issues.
  • Support large-scale production environments with a strong focus on reliability, automation, and operational excellence.
  • Work with financial services and trading platform environments when applicable.

Required Skills

  • Site Reliability Engineering (SRE)
  • SLOs and SLIs
  • Dynatrace
  • Python or Perl
  • AWS or Azure
  • DevOps
  • Strong Python development experience
  • Hands-on Django and REST API development
  • Strong MySQL and database skills
  • Deep Linux administration and troubleshooting experience
  • Infrastructure engineering and automation
  • Ansible
  • CI/CD pipelines
  • Observability and monitoring
  • OpenTelemetry
  • Infrastructure-as-Code
  • Automation frameworks
  • Experience working in large-scale production environments

Preferred Qualifications

  • Financial Services, Capital Markets, or Trading Platform experience
  • Experience with cloud migration and modernization
  • Experience with performance engineering and application monitoring
  • Experience with distributed systems and cloud-native technologies
  • Strong production support, incident management, and troubleshooting experience

Similar jobs