Why This Role Stands Out
This role offers an exceptional opportunity to push the boundaries of machine learning performance in a fast-paced, high-impact environment, with the potential for significant career growth within a renowned financial technology firm. You'll thrive here if you possess a deep understanding of low-level systems, a passion for optimization, and a curious mind eager to tackle complex challenges in real-time trading. Apply now to join a collaborative team and make a tangible difference.
Quick Overview
Job Description
We are looking for an engineer with experience in low-level systems programming and optimisation to join our growing ML team.
Machine learning is a critical pillar of Jane Street's global business. Our ever-evolving trading environment serves as a unique, rapid-feedback platform for ML experimentation, allowing us to incorporate new ideas with relatively little friction.
Your part here is optimising the performance of our models – both training and inference. We care about efficient large-scale training, low-latency inference in real-time systems and high-throughput inference in research. Part of this is improving straightforward CUDA, but the interesting part needs a whole-systems approach, including storage systems, networking and host- and GPU-level considerations. Zooming in, we also want to ensure our platform makes sense even at the lowest level – is all that throughput actually goodput? Does loading that vector from the L2 cache really take that long?
If you’ve never thought about a career in finance, you’re in good company. Many of us were in the same position before working here. If you have a curious mind and a passion for solving interesting problems, we have a feeling you’ll fit right in.
There’s no fixed set of skills, but here are some of the things we’re looking for:
- An understanding of modern ML techniques and toolsets
- The experience and systems knowledge required to debug a training run’s performance end to end
- Low-level GPU knowledge of PTX, SASS, warps, cooperative groups, Tensor Cores and the memory hierarchy
- Debugging and optimisation experience using tools like CUDA GDB, NSight Systems, NSight Computesight-systems and nsight-compute
- Library knowledge of Triton, CUTLASS, CUB, Thrust, cuDNN and cuBLAS
- Intuition about the latency and throughput characteristics of CUDA graph launch, tensor core arithmetic, warp-level synchronization and asynchronous memory loads
- Background in Infiniband, RoCE, GPUDirect, PXN, rail optimisation and NVLink, and how to use these networking technologies to link up GPU clusters
- An understanding of the collective algorithms supporting distributed GPU training in NCCL or MPI
- An inventive approach and the willingness to ask hard questions about whether we're taking the right approaches and using the right tools
- Fluent in English
If you're a recruiting agency and want to partner with us, please reach out to agency-partnerships@janestreet.com.
Similar jobs
- PI
Staff Machine Learning Engineer, Content Visual AI
NewAuto ApplyPinterest
San Francisco🇺🇸Remote4 hours agoMachine LearningHiveLLMTechnology - HE
Inference Optimization Engineer
NewAuto ApplyHedra
San Francisco🇺🇸Hybrid6 hours agoCUDADeep LearningC+++2Engineering - ST
Machine Learning Engineer, Radar
NewAuto ApplyStripe
Seattle🇺🇸Hybrid21 hours agoSQLDeep LearningPyTorch+1Technology - TR
Machine Learning Engineer, Applied
NewAuto ApplyTracelabs
United States🇺🇸Remote21 hours agoRoboticsComputer VisionDeep Learning+3Technology - CO
Staff Machine Learning Engineer, Personalization
NewAuto ApplyCoupang
Mountain View🇺🇸$152k - $277k/yrHybrid23 hours agoAWSMLflowMachine Learning+11Technology - RM
Member of Technical Staff, ML Platform
NewAuto ApplyRunway Ml
Remote🇺🇸RemoteYesterdayRoboticsKubernetesLESS+3 - DE
Staff Inference Engineer
NewAuto ApplyDesignworkstalent
Bellevue🇺🇸HybridYesterdayEngineering - BU
Staff Machine Learning / Operations Research Engineer
NewAuto ApplyBurq, Inc.
United States🇺🇸Remote22 hours agoMLOpsDeep LearningForecasting+2Technology - OP
Senior AI/ML Test and Evaluation Engineer
NewAuto ApplyOpenTeams
United States - Remote OR Hybrid🇺🇸$145k - $250k/yrRemoteYesterdayExpressMachine LearningNumPy+6Technology - CL
Machine Learning Engineer (AI/ML)
NewAuto ApplyClickhouse
New York🇺🇸RemoteYesterdayGCPAWSMachine Learning+5Technology - WO
AI/ML Engineer
NewAuto ApplyWoolpert
Remote - United States🇺🇸$118.2k - $147.8k/yrRemote2 days agoGCPAgileArticulate+7Technology - PI
Director, Machine Learning Engineering, Ads Quality
NewAuto ApplyPinterest
Palo Alto🇺🇸$314.6k - $550.5k/yrHybridYesterdayMachine LearningTechnology