Haystack
← Back to Jobs
Technology
TS

System Engineer 3

Talent Software Services, IncRedmond, WA🇺🇸United StatesPosted 18 Aug 2026

Quick Overview

Work Type
Hybrid
Level
Mid Senior

Job Description

Typical Day in the Role

Purpose of the Team: The purpose of this team is to manage Azure resources.and have to maintain compliance on some of those, on, well, on all of those resources to just make sure that, hey, we''re aligned with the latest security, either patching or managing of resources, facilitating with upgrades of those pieces of material.

Candidate Requirements
•Best vs. Average: The ideal resume would contain working knowledge of Ansible and playbooks

Summary:
This is a site reliability engineering role in the Service Health Platforms team, responsible for the operation of enterprise network observability platforms that consist of commercially available platforms and internally developed systems that ingest network telemetry. The data ingested by these systems provide the foundation for tools used by our network engineers to operate more efficiently in the context of scale.

Job Responsibilities:
• Administer and operate Windows and Linux VMs hosted in Azure to ensure they remain complaint with security and configuration standards.
• Maintain network observability platforms including syslog-ng, and trapd by performing upgrades, patching, capacity planning, authoring rules using regular expressions.
• Provide additional support for IBM SevOne Network Performance Manager and Broadcom AppNeta observability platforms
• Manage assigned projects and program components to deliver services in accordance with established objectives
• Identify and automate tasks with Bash, Powershell, or Python
• Troubleshoot complex network observability configurations, software applications, operating systems through regular maintenance
• Deploy, configure, and manage cloud services on platforms like Azure, ensuring scalability, reliability, and cost-effectiveness.
• Implement DevOps practices and tools, including CI/CD pipelines and infrastructure as code, to automate and streamline development and deployment processes.
• Perform BCDR failover testing where appropriate
• Participate in on-call rotation (DRI)

Required Qualifications:
• Bachelor''s degree in a technical field such as computer science, computer engineering or related field; or equivalent experience
• 5-7 years enterprise experience in an IT systems, network, or site reliability engineering role
• Strong understanding of network observability, including SNMP, SNMP Traps, NetFlow, gNMI
• Hands-on experience with cloud platforms such as Azure or similar cloud platforms
• Strong understanding of enterprise networking protocols
• Experience with system capacity and planning, as well as functional configuration and audit
• Experience with system planning and capacity tools and analyses

Preferred Qualifications:
• Proficiency in automating tasks using Bash, Powershell, Python.
• Proficient with regular expressions
• Experience with IBM SevOne and/or Broadcom AppNeta or similar observability platforms.
• Experienced with various source control platforms
• Working knowledge of Ansible and playbooks
• Intermediate knowledge of data retrieval languages such as KQL and T-SQL
• Experience and exposure on commercially available AI platforms

What are unique selling points that would get candidates interested in your role over another?
The candidate would get an opportunity to work in a very large enterprise environment and gain expertise to tie network telemetry to AI workflows for performance, capacity management, and security purposes.

What is the ideal background of a candidate for this role?
The ideal background for the role would to have deep understanding of network monitoring systems and management protocols. Technical proficiency with enterprise routing and switching protocols. Deep knowledge running syslog-ng and trapd on Linux. Exposure to management platforms such as IBM SevOne Network Performance Manager and Broadcom AppNeta.

Explain a typical day in the role.
The candidate will be responsible for operating network observability platforms with a primary focus on syslog-ng and trapd running on Linux. Planning and executing security/operating system/application patching. Investigating automated alerts generated from monitoring and customer reported incidents related to the network observability platforms. Assisting network and security engineers in identification of traffic patterns and resource utilization. Attending standup with other engineers in the team to review and prioritize work for on-time delivery.

How will contractor performance be measured?
The contractor will be measured by compliance outcomes in internal systems such as S360, incident resolution, and ability to deliver timely and effective outcomes for projects.

Top 3 Must-Have HARD Skills & years of experience for each
Syslog-NG (3 years)
Linux sysadmin (5 years)
Network Engineering (3 years)

Where is the work able to be performed?
Remote

Skills

SQL
T-SQL
Ansible
Azure
Bash
PowerShell
Python

Similar jobs