← Back to Jobs
Technology
AP
Machine Learning Engineer, Foundation Model Services
Apple, Inc.Santa Clara, CA🇺🇸United StatesPosted 12 Aug 2026
Quick Overview
Work Type
Hybrid
Level
Mid Senior
Job Description
Do you feel you think differently, you are eager to break status quo, are bold and ambitious, aren't afraid to take risks and are passionate to build the best of class technology. If yes, what better place to be at and do this than Apple? At Apple, "we think different, we push the boundaries of computing and intelligence. We build products that bring smile to people's face".
Foundation Model Services team, within Machine Learning Platform Technologies organization is the back-bone of Apple Intelligence. It builds frameworks, services and tools that power the largest Apple foundation models on servers. Our Infrastructure powers a wide gamut of services at Apple including Apple Search, Apple Music, AppleTV, AppStore, iMessages, Photos & Camera, Spotlight, Safari, Siri and upcoming ever exciting Apple products serving millions of queries every day with incredible low latencies, drawing every ounce of compute from our hardware. As part of this group, you will get a chance to bring Intelligence to billions of users across the world. You will have an opportunity to make a difference in life of people. You will have a chance to work on optimizing billions of parameter language and vision and speech models using state of the art technologies and make it run at scale of Apple.
Description
* Work closely with product teams to build production grade solutions to launch models serving millions of customers in real time.
* Work along side Foundation Model Research team to prototype and develop inference for cutting edge model architectures.
* Build tools to understand bottlenecks in Inference for different hardwares and use cases.
Minimum Qualifications
5 year+ industry experience in ML technologies (LLMs, Machine Learning, NLP, Information Retrieval, Statistics).
Experience with high throughput services particularly at supercomputing scale.
Proficient with running applications on Cloud (AWS / Azure or equivalent) using Kubernetes, Docker etc.
Proficient in building and maintaining systems written in modern languages (eg: Golang, python)
Bachelor's degree or higher in Computer Science or related technical field.
Preferred Qualifications
Familiar with one of the popular ML Frameworks like Pytorch, Tensorflow.
Familiar with fundamental Deep Learning architectures such as Transformers, Encoder/Decoder models.
Familiarity with Nvidia TensorRT-LLM, vLLLM, DeepSpeed, Nvidia Triton Server etc.
Foundation Model Services team, within Machine Learning Platform Technologies organization is the back-bone of Apple Intelligence. It builds frameworks, services and tools that power the largest Apple foundation models on servers. Our Infrastructure powers a wide gamut of services at Apple including Apple Search, Apple Music, AppleTV, AppStore, iMessages, Photos & Camera, Spotlight, Safari, Siri and upcoming ever exciting Apple products serving millions of queries every day with incredible low latencies, drawing every ounce of compute from our hardware. As part of this group, you will get a chance to bring Intelligence to billions of users across the world. You will have an opportunity to make a difference in life of people. You will have a chance to work on optimizing billions of parameter language and vision and speech models using state of the art technologies and make it run at scale of Apple.
Description
* Work closely with product teams to build production grade solutions to launch models serving millions of customers in real time.
* Work along side Foundation Model Research team to prototype and develop inference for cutting edge model architectures.
* Build tools to understand bottlenecks in Inference for different hardwares and use cases.
Minimum Qualifications
5 year+ industry experience in ML technologies (LLMs, Machine Learning, NLP, Information Retrieval, Statistics).
Experience with high throughput services particularly at supercomputing scale.
Proficient with running applications on Cloud (AWS / Azure or equivalent) using Kubernetes, Docker etc.
Proficient in building and maintaining systems written in modern languages (eg: Golang, python)
Bachelor's degree or higher in Computer Science or related technical field.
Preferred Qualifications
Familiar with one of the popular ML Frameworks like Pytorch, Tensorflow.
Familiar with fundamental Deep Learning architectures such as Transformers, Encoder/Decoder models.
Familiarity with Nvidia TensorRT-LLM, vLLLM, DeepSpeed, Nvidia Triton Server etc.
Skills
Docker
AWS
Machine Learning
NLP
Azure
Deep Learning
Go
Kubernetes
LLM
PyTorch
Python
TensorFlow
Similar jobs
Staff Machine Learning Engineer - Data
Kodiak · San Francisco, United States
1 hour ago$200k - $265k/yrStaff Machine Learning Engineer - Search
Warner Bros. Discovery · San Francisco, United States
1 hour ago$192.6k - $357.6k/yrMachine Learning Engineer
Unity Microelectronics Inc. · Mountain View, United States
1 hour ago$117k - $152k/yrPrincipal Machine Learning Engineer
Unity Microelectronics Inc. · Mountain View, United States
1 hour ago$278.1k - $417.1k/yrSenior/Staff Machine Learning Engineer, Data Infrastructure
Unity Microelectronics Inc. · Mountain View, United States
1 hour ago$200.4k - $260.5k/yrSenior Machine Learning Engineer - Foundation Model
XPENG · Santa Clara, United States
1 hour ago$174.7k - $295.7k/yr