Haystack
← Back to Jobs
Technology
BC

Site Reliability Engineer III

BCforwardPennington, NJ🇺🇸United StatesPosted 12 Sept 2026

Why This Role Stands Out

This role offers exciting opportunities to enhance critical trading and business services through innovative SRE practices and automation, perfect for a mid-senior engineer eager to drive reliability and efficiency. You'll thrive by leveraging your skills in observability tools and automation to build robust solutions, enjoying a flexible hybrid work arrangement. Apply now to contribute to a leading IT consulting firm and advance your career in a dynamic tech environment.

Quick Overview

Salary
$74/hr
Seniority
Mid Senior
Work mode
Hybrid
Location
Pennington, NJ, United States
Posted
16 hours ago
MongoDBMySQLNode.jsShellSplunkAnsibleGrafanaKubernetesPrometheusPythonTerraform

Job Description



Job Title: Site Reliability Engineer III


Location: Pennington, NJ


Duration: Contract - 7 months


Pay Range: $73.67/hr (W2)


Job ID: 409685



About BCforward


BCforward is a leading global IT consulting and workforce solutions firm providing services and support to Fortune 500 and government clients. Founded in 1998, BCforward has grown with our customers needs into a full-service business solutions provider. With delivery centers and offices across North America and India, we take pride in building long-term relationships and delivering excellence through innovation, collaboration, and integrity.



Job Description


We are seeking a Site Reliability Engineer III to join our dynamic team within Global Markets. The successful candidate will design, build, and support tooling, automation frameworks, and observability capabilities that enable reliable operation of critical trading and business services at scale. The role will drive SRE best practices, improve operational workflows, enhance platform observability, and develop solutions that reduce manual effort while increasing service reliability and resilience.


Work Arrangement: Onsite 4 days per week for the first six months, then 3 days onsite and 2 days remote.


Reference: BOA-VidVoiceEng-I-V-1


Responsibilities:



  • Design, develop, and maintain SRE tooling, automation frameworks, and engineering utilities.

  • Implement and optimize observability using platforms such as Dynatrace, Splunk, Grafana, and Prometheus.

  • Define and operationalize SLIs, SLOs, and error budgets to improve service reliability.

  • Build monitoring, alerting, and operational dashboards to support production excellence.

  • Lead and support incident management, root cause analysis, and continuous improvement actions.

  • Integrate systems and services through APIs and service interfaces to streamline workflows.

  • Contribute to CI/CD pipeline design and deployment automation for reliable releases.

  • Partner with application, infrastructure, and production support teams to embed SRE practices.

  • Conduct reliability engineering, resiliency assessments, and operational readiness reviews.

  • Support cloud migration and modernization programs with reliable architectures.


Required Skills & Qualifications:



  • Strong Linux administration and troubleshooting skills.

  • Proficiency in Python and Shell scripting or similar automation languages.

  • Hands-on experience with observability platforms such as Dynatrace, Splunk, Grafana, and Prometheus.

  • Experience implementing monitoring, alerting, and operational dashboards.

  • Incident management, production support, and root cause analysis expertise.

  • Experience integrating systems via APIs and service interfaces.

  • Knowledge of cloud platforms and distributed systems architecture.

  • Experience with CI/CD pipelines and deployment automation.

  • Understanding of database technologies such as MySQL and MongoDB, and familiarity with Node.js.

  • Strong analytical, problem-solving, and debugging skills.

  • Experience in a Site Reliability Engineering organization with cross-team collaboration.

  • Hands-on experience defining and implementing SLIs, SLOs, and error budgets.


Preferred Skills:



  • Kubernetes and OpenShift experience.

  • Infrastructure as Code using Terraform or Ansible.

  • Development of self-service platforms and engineering productivity tools.

  • AIOps, intelligent alerting, and automated incident triage solutions.

  • Knowledge of observability standards including OpenTelemetry.

  • Experience in large-scale, high-availability financial services environments.


Why BCforward?


At BCforward, we believe in advancing lives and careers. When you join our team, you gain access to:



  • Competitive compensation and benefits.

  • Opportunities for growth with global clients.

  • A supportive, inclusive culture that values innovation and people.

  • Exposure to cutting-edge technologies and projects.



About Our Commitment


BCforward is an equal opportunity employer. We value diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, sexual orientation, gender identity, national origin, age, disability, or veteran status.


Interested? Apply Now!


If this sounds like the right opportunity for you, please apply with your most recent resume.


Similar jobs