Haystack
← Back to Jobs
Technology
OT

Machine Learning Inference Engineer

Oscar TechnologySan Francisco, CA🇺🇸United StatesPosted 25 Aug 2026

Why This Role Stands Out

This hybrid role offers a competitive salary of $200K-$250K plus equity at a rapidly growing AI startup, allowing you to directly impact the performance and scalability of next-generation AI applications. You'll thrive here if you're a mid-senior engineer passionate about optimizing ML inference infrastructure, solving complex performance challenges, and building production-ready AI systems. Apply now to join a collaborative team and accelerate your career in cutting-edge AI development.

Quick Overview

Salary
$200k - $250k/yr
Seniority
Mid Senior
Work mode
Hybrid
Location
San Francisco, CA, United States
Posted
1 week ago
MicroservicesMachine LearningPyTorchPython

Job Description



Title: ML Inference Engineer


Location: San Francisco, CA


Salary: $250k base + equity


An AI Unicorn startup is hiring a Senior Machine Learning Inference Engineer for a full-time role.


You will be responsible for improving efficiency for AI-native infrastructure powered by generative and multimodal models. The ideal candidate has over 3 years of professional experience and a strong understanding of GPU infrastructure, Python, and PyTorch. This is a highly autonomous role with significant ownership across inference systems and model performance in production.


This role is hybrid in San Francisco Bay Area and offers full benefits and equity.


Experience:



  • Building AI applications at scale from the ground up

  • Strong understanding of GPU infrastructure including Triton, TensorRT, or vLLM frameworks
    Hands-on experience with Python and PyTorch

  • Building model-serving Microservices

  • Diffusion and Multimodal model experience is a plus


Benefits:



  • Competitive base salary

  • Equity

  • $401k matching

  • Medical coverage



Oscar Associates Limited (US) is acting as an Employment Agency in relation to this vacancy.

Similar jobs