Quick Overview
Job Description
Data Center Operations — Weekday Overnight
On-site in Pittsburgh, Pennsylvania | Monday–Friday, 12:00 a.m.–8:00 a.m. Eastern
Each shift begins at midnight on the listed day. The Monday shift therefore begins late Sunday night, and the Friday shift ends Friday morning.
Who We Are
Teraswitch is on a mission to provide the highest-performance, lowest-latency bare metal servers in the world. With 20 data center locations around the world, Teraswitch has served thousands of customers across 185 countries.
Founded by Brendan Mannella and headquartered in Pittsburgh, Pennsylvania, Teraswitch is a privately held infrastructure company operating a growing global bare metal platform.
What to Expect
As a member of Data Center Operations, you will own infrastructure incidents, customer requests, maintenance activities, and deployment projects from intake through verified completion.
This is primarily a remote infrastructure-operations role—not a primarily hands-on data-center technician position. Most work is performed through monitoring systems, tickets, server-management interfaces, technical documentation, and coordination with customers, internal engineering teams, colocation facilities, vendors, contractors, and remote-hands providers.
Approximately 10% of the role may involve physical work including server assembly, component replacement, cabling, rack-and-stack work, and equipment maintenance.
This role requires someone who can investigate problems independently, communicate clearly with technical and nontechnical stakeholders, create precise technical procedures, manage multiple concurrent priorities, and remain accountable until work is completed and validated.
Responsibilities include:
Own infrastructure incidents, customer requests, maintenance activities, deployments, and other operational projects from intake through verified completion.
Monitor infrastructure alerts, support requests, maintenance activities, and active operational issues.
Investigate reported server, network, power, cooling, connectivity, deployment, and monitoring problems.
Gather and interpret information from logs, monitoring platforms, tickets, management interfaces, internal systems, customers, and on-site technicians.
Perform initial incident triage, take safe and documented corrective actions, and escalate service-affecting issues when appropriate.
Communicate directly and professionally with customers during incidents, technical investigations, maintenance events, and service requests.
Write precise Methods of Procedure, implementation plans, validation steps, rollback instructions, runbooks, and remote-hands directions.
Translate technical requirements into clear, executable instructions for colocation technicians, contractors, vendors, and internal teams.
Coordinate and direct remote-hands technicians at Teraswitch data center locations around the world.
Remain accountable for remote work by confirming the correct equipment, reviewing evidence, tracking progress, validating results, and ensuring documentation is complete.
Coordinate scheduled installations, migrations, maintenance events, hardware replacements, shipments, and other operational projects.
Work with facility personnel to address power, cooling, cross-connect, cabling, access, and other facility-provided services.
Coordinate vendor warranty technicians and ensure hardware issues are resolved promptly and correctly.
Maintain clear ticket updates, incident timelines, project notes, escalation records, and actionable shift handoffs.
Collaborate with Software, Network, Systems, Facilities, Logistics, and other internal teams.
Maintain accurate records of infrastructure, hardware changes, completed work, and physical inventory.
Identify recurring problems and improve documentation, procedures, tooling, and operational workflows.
Execute documented server-management, recovery, deployment, and maintenance procedures.
Perform occasional hands-on server assembly, rack-and-stack work, cabling, diagnostics, upgrades, and component replacement in Pittsburgh.
Follow established security, change-management, escalation, validation, and access-control procedures.
What Our New Team Member Will Need
Three or more years of experience in data center operations, infrastructure support, systems support, technical operations, or a similar role.
Demonstrated ability to write detailed technical procedures that another technician can execute without additional interpretation.
Experience troubleshooting infrastructure remotely using logs, monitoring platforms, tickets, command-line tools, server-management interfaces, or information supplied by on-site personnel.
Experience independently owning incidents, service requests, maintenance activities, or technical projects through completion.
Strong written communication and the ability to provide clear updates to customers, vendors, contractors, and internal technical teams.
Ability to prioritize multiple concurrent customer, infrastructure, and operational issues.
Sound judgment regarding service impact, change control, escalation, rollback, and validation.
Working knowledge of server hardware, network equipment, structured cabling, and data center facility services.
Ability to distinguish between gathering information, taking corrective action, escalating an issue, and transferring ownership.
Ability to follow technical procedures precisely and verify that completed work produced the intended result.
Strong troubleshooting, organization, and time-management skills.
Ability to safely lift, move, rack, and work around data center equipment when physical work is required.
Availability to consistently work the stated weekday overnight schedule on-site in Pittsburgh.
Helpful Experience
Writing MOPs, SOPs, runbooks, implementation plans, rollback procedures, and escalation documentation.
Coordinating remote-hands providers, contractors, vendors, and colocation facility personnel.
Linux administration and troubleshooting.
Server management through IPMI, Redfish, iDRAC, iLO, or other BMC platforms.
Infrastructure monitoring and alert-management platforms.
Ticketing, documentation, inventory, and infrastructure-management systems.
Basic networking concepts, including IP addressing, VLANs, interface bonding, and physical connectivity.
Server diagnostics, component replacement, firmware management, and hardware compatibility.
Hardware inventory and spare-parts management.
Experience supporting bare metal, hosting, cloud, network, or large-scale distributed infrastructure.
Familiarity with APIs or API endpoints used to retrieve and validate operational data.
Compensation and Benefits
Along with competitive pay, full-time Teraswitch employees are eligible for the following benefits beginning on their first day of employment:
Health, dental, and vision insurance.
401(k) with company profit sharing.
Flexible paid time off.
Company holiday benefits.
Success in This Role
Success means operational issues receive prompt attention, technical evidence is gathered before potentially disruptive action is taken, customers and internal teams receive clear updates, and remote technicians receive instructions they can execute without ambiguity.
Work is not considered complete merely because it was escalated or assigned to another party. Success means the employee retains ownership, coordinates the required people, tracks progress, validates the result, updates the relevant documentation, and provides a complete and actionable handoff when work continues into another shift.