Network Operations Engineer (5+ years)
Why This Role Stands Out
This role offers an exciting opportunity to be at the forefront of autonomous vehicle technology, directly impacting the reliability of live operations and developing your skills in a mission-critical environment. You'll thrive here if you possess a strong technical curiosity, excellent troubleshooting abilities, and the composure to excel in high-pressure situations. Join a dynamic team dedicated to innovation and operational excellence.
Quick Overview
Job Description
We are helping an on-demand, autonomous ride-hailing company find Network Operations Engineers to monitor and support the critical networks, cloud infrastructure, databases, and authentication systems that keep live autonomous vehicle operations running reliably.
In this role, you'll monitor critical infrastructure, triage alerts, respond to incidents, and escalate complex issues to specialized engineering teams. You'll also partner closely with Site Reliability Engineering (SRE) to identify recurring issues and turn operational insights into long-term reliability improvements.
The ideal candidate has experience working in a NOC, TechOps, or similar mission-critical operations environment and is comfortable troubleshooting infrastructure issues in real time. You are highly responsive, technically curious, and able to communicate clearly and remain composed during high-pressure incidents.
As a Network Operations Engineer, you'll:
- Proactively monitor critical IT services, networks, and infrastructure supporting live autonomous vehicle operations in a 24/7 environment.
- Monitor the real-time health of cloud infrastructure, databases, DNS, authentication services, and other systems required for continuous operations.
- Acknowledge, investigate, and triage system alerts, serving as a primary responder for incidents affecting live services.
- Troubleshoot incidents and participate in escalation calls to support investigation, resolution, and root cause identification.
- Follow established operational procedures and escalation paths to route complex issues to the appropriate engineering teams.
- Partner with SRE and engineering teams to identify recurring issues and improve system reliability.
- Participate in incident retrospectives and recommend improvements that reduce outages and strengthen operational processes.
- Document incidents, troubleshooting activities, resolutions, and escalation details accurately.
- Prepare clear shift handoffs and service status reports to maintain continuity across 24/7 operations.
- Experience: 5+ years in a NOC, SOC, TechOps, or similar structured operations environment, with a strong understanding of incident management lifecycles and SLAs.
- Infrastructure Monitoring: Proven experience monitoring enterprise networks, cloud infrastructure (AWS or GCP), and critical databases.
- Core Services: Strong understanding of TCP/IP, DNS, authentication services EntraID, OpenAuth, SSL/TLS and general networking.
- Tooling: Proficiency with modern monitoring and observability platforms (e.g., Grafana, Datadog, Kentik, Logic Monitor, or similar alert management systems).
- Communication: Excellent written and verbal communication skills, with the ability to convey critical technical issues clearly during high-pressure situations.
- Availability: Willing and able to work 100% on-site in Scottsdale, AZ on an assigned shift that includes at least one weekend day (Saturday or Sunday) every week, including holiday coverage as scheduled.
Preferred Qualifications:
- Experience with ticketing and incident management systems such as Incident.io, Pagerduty, Opsgenie, ServiceNow and JIRA
- Basic scripting for operational tasks (Python, Bash) Linux fundamentals ITIL or comparable incident/service management framework familiarity
- CompTIA Network+ and Security+ certifications Cisco Certified Network Associate (CCNA) or similar vendor-specific networking certifications (e.g., Juniper JNCIA)
401K
Life Insurance
Health Insurance
Skills
Similar jobs
Software Engineer (TGF)
Air Traffic Engineering Co. LLC · Egg Harbor Township, United States
6 minutes agoSystems Reliability Engineer II
Nutanix · Durham, United States
7 minutes ago$142.8k/yrSenior Systems Engineer, Air Vehicle Software
Anduril Industries · Costa Mesa, United States
7 minutes ago$166k - $220k/yrSenior UI/UX Developer / Front-End Developer
BCforward · Jacksonville, United States
8 minutes ago$65/hrMid Cloud DevOps Engineer
Booz Allen Hamilton · McLean, United States
8 minutes ago$62k - $141k/yrSenior Systems/ Infrastructure Engineer with Security Clearance
Amentum · McLean, United States
8 minutes ago$140k - $160k/yr