Why This Role Stands Out
This HPC Systems Engineer role offers a unique opportunity to optimize cutting-edge research infrastructure and drive significant impact within a reputable trading firm. You'll thrive here if you're a proactive problem-solver with deep Linux and HPC expertise, eager to lead complex projects and develop innovative solutions. Apply now to elevate your career in a dynamic, collaborative environment.
Quick Overview
Job Description
Radix Trading is seeking a seasoned HPC Systems Engineer with a passion for Linux, HPC systems and research infrastructure to join us. This role requires the ability to quickly troubleshoot and resolve technical issues, navigate networking and system alerts, and collaborate with internal teams and external partners. The ideal candidate is a self-starter that has experience leading projects end-to-end, fluency in scripting languages, and a proven track record of solving complex technical problems.
Job Functions:
- Own monitoring and performance tuning across HPC applications, SLURM job scheduler, networking, storage, and hardware to optimize workload efficiency.
- Perform root-cause analysis to develop sustainable automated solutions.
- Support and troubleshoot alerts and errors with savvy communication skills to interact with external entities and internal counterparts.
- Keep up communication with new and existing external vendors on technical infrastructure work.
- Keep up communication with new and existing external business interfaces --- day-to-day operation, planning connectivity, and exchange upgrades.
- Oversee communication with others internally - project management across different functions to plan exchange upgrades, hardware refreshes, and other improvements/rollouts.
- Build out and support research, trading, and enterprise infrastructure.
Qualifications:
- 5+ years of HPC system administration/architecture including RHEL/CentOS/Rocky Systems
- Experience with job scheduler and resource management tools such as SLURM, Moab, or Torque
- Knowledge of network storage systems such as DDN, IBM SpectrumScale, NetApp, Weka, or Vast
- Knowledge of parallel file systems such as Lustre or Spectrum Scale (GPFS)
- Experience working with InfiniBand and high-speed Ethernet
- Hands-on experience with configuration management tools such as xCAT, Ansible, Salt, and Terraform
- Exposure to bare metal provisioning — including DHCP, DNS, PXE Boot
- Competence in Python, able to edit and create scripts
Similar jobs
- SA
Senior Systems Administrator
NewSaalex
Ridgecrest, California🇺🇸$38 - $40/hrOn-site12 hours agoDockerAWSLoad Balancing+18Technology - SA
Journeyman Systems Administrator
NewSaalex
Ridgecrest, California🇺🇸On-site12 hours agoDockerAWSLoad Balancing+18Technology - CA
Systems Administrator
NewCAI
United States🇺🇸$47 - $50/hrOn-site48 minutes agoDockerGCPOracle+12Technology - OR
System Administrator 1-IT
NewOracle
Kansas City, Missouri🇺🇸$26 - $50/hrOn-site48 minutes agoOracleTechnology - NE
IT System Administrator
NewNextern
Maple Grove, Minnesota🇺🇸On-site9 hours agoTCP/IPActive DirectoryAzure+3Technology - GE
Senior Workfront System Administrator
NewGenesis10
Cleveland, OH🇺🇸$33 - $43/hrHybridYesterdayTechnology