Quick Overview
Job Description
Youl'll be designing, deploying, andoperatingthe switch fabrics and interconnects that power large-scale training and inference workloads.
16th September, 2026About the Role
One of our clients are building out their network backbone that connects and orchestrates thousands of GPUs at scale.
You'llbe joining a team working withcutting-edgeNvidia hardware and high-performance networking technologies that few localorganisationsoperateat this scale.
The role:
- Design, build, andoperatehigh-performance network fabrics supporting large-scale GPU clusters
- Deploy and tune switch fabrics for low-latency, high-throughput AI/ML workloads
- Work closely with infrastructure and platform teams tooptimisenetwork performance for distributed training and inference
- Troubleshoot and resolve complex network issues across HPC and GPU-dense environments
- Plan capacity and topology for ongoing GPU infrastructure expansion
- Partner with vendors (including Nvidia) on hardware qualification, firmware, and fabric design
- Contribute to standards and best practices for network architecture as the environment scales
About the Role
One of our clients are building out their network backbone that connects and orchestrates thousands of GPUs at scale.
You'llbe joining a team working withcutting-edgeNvidia hardware and high-performance networking technologies that few localorganisationsoperateat this scale.
The role:
- Design, build, andoperatehigh-performance network fabrics supporting large-scale GPU clusters
- Deploy and tune switch fabrics for low-latency, high-throughput AI/ML workloads
- Work closely with infrastructure and platform teams tooptimisenetwork performance for distributed training and inference
- Troubleshoot and resolve complex network issues across HPC and GPU-dense environments
- Plan capacity and topology for ongoing GPU infrastructure expansion
- Partner with vendors (including Nvidia) on hardware qualification, firmware, and fabric design
- Contribute to standards and best practices for network architecture as the environment scales
Experience needed:
- Hands-on experience withhigh performancecomputing (HPC) networking environments
- Strong understanding of GPU infrastructure and the networking demands of AI/ML workloads
- Experience with switch fabric design, deployment, and operations at scale
- Familiarity with Nvidia networking technologies (e.g.InfiniBand, Spectrum-X,NVLink/NVSwitchecosystems)
- We'llalso consider candidates without direct HPC/AI experience if they bring:
- Large-scale network engineering experience from ahyperscaleror major cloud provider (e.g.AWS, Azure, Google Cloud) or equivalent big-tech environment
- Proven experience designing oroperatinglarge, complex switch fabrics in production - this background is scarce in the Australian market and highly valued
- A track recordof operating at scale (thousands of nodes/ports) rather than traditional enterprise networking
Similar jobs
- BJ
Mobile Engineer - AI Finance Agent
NewBjak
Sydney, New South Wales🇦🇺Hybrid2 hours agoSQLAndroid SDKJetpack Compose+5Technology - CB
Staff Data Engineer (AWS, Java & Real-Time Data)
NewCommonwealth Bank of Australia
Melbourne🇦🇺Hybrid5 hours agoSQLAWSFlink+9Technology - XS
Senior Full Stack Developer
NewXPT Software Australia Pty
Sydney, New South Wales🇦🇺Hybrid5 hours agoExpressMaterial UIMicroservices+16Technology - SA
Salesforce Developer
NewSlater and Gordon
Melbourne, Victoria🇦🇺Hybrid5 hours agoSOAPAgileCSS+4Technology - ON
Applied AI Engineer
NewOnebyZero .
Sydney, New South Wales🇦🇺Hybrid5 hours agoDockerAWSMachine Learning+3Technology - SS
IT Support Engineer (Service Desk)
NewSteadfast Solutions
Milton, Queensland🇦🇺A$80k - A$90k/yrHybrid5 hours agoMFATCP/IPAzure+1Technology