← Back to Jobs
Technology
Data Scientist - Survey Design, Data Annotation, and Machine Learning Evaluation
Apple, Inc.Cupertino, CA🇺🇸United StatesPosted 12 Aug 2026
Quick Overview
Work Type
Hybrid
Level
Mid Senior
Job Description
Apple is where individual imaginations gather together, committing to the values that lead to
great work. Every new product we build, service we create, or experience we deliver is the
result of us making each other's ideas stronger. The diversity of our people and their thinking
inspires the innovation that runs through everything we do. When we bring everybody in, we
can do the best work of our lives. Here, you'll do more than join something - you'll add
something.
Description
The Special Projects team at Apple is developing novel user-facing conversational features that
leverage the multimodal capabilities of state-of-the-art foundation models. As part of this
process, we generate real-world and simulated data, gather human data annotations, analyze
the results, and use them to build and evaluate Large Language Model judges. We are looking
for a skilled Data Scientist to join our Machine Learning Evaluations teams. This person will
work closely with ML Engineers to manage and analyze our human and automated data
annotation processes, and to develop, test, and refine LLM judges for generative AI model
evaluation. A successful candidate is experienced in survey design, data annotation, LLM
prompt engineering and prompt optimization, and has strong statistical analysis skills.
Minimum Qualifications
BA or Master's degree in Data Science, Statistics, or a quantitative social science field
2+ years of hands-on experience working in survey design and human data annotation
Proficiency in Python
Excellent communication skills
Preferred Qualifications
PhD in Data Science, Statistics, or a quantitative social science field
Hands-on industry experience with product-focused statistical analysis
Experience working with large-scale multimodal data and data-annotation pipelines
Experience with LLM prompt engineering & prompt optimization
Experience with LLM auto-judges for generative AI model evaluation
A track record of publications or technical presentations in Data Science or a related field
Excellent at cross-functional collaboration
great work. Every new product we build, service we create, or experience we deliver is the
result of us making each other's ideas stronger. The diversity of our people and their thinking
inspires the innovation that runs through everything we do. When we bring everybody in, we
can do the best work of our lives. Here, you'll do more than join something - you'll add
something.
Description
The Special Projects team at Apple is developing novel user-facing conversational features that
leverage the multimodal capabilities of state-of-the-art foundation models. As part of this
process, we generate real-world and simulated data, gather human data annotations, analyze
the results, and use them to build and evaluate Large Language Model judges. We are looking
for a skilled Data Scientist to join our Machine Learning Evaluations teams. This person will
work closely with ML Engineers to manage and analyze our human and automated data
annotation processes, and to develop, test, and refine LLM judges for generative AI model
evaluation. A successful candidate is experienced in survey design, data annotation, LLM
prompt engineering and prompt optimization, and has strong statistical analysis skills.
Minimum Qualifications
BA or Master's degree in Data Science, Statistics, or a quantitative social science field
2+ years of hands-on experience working in survey design and human data annotation
Proficiency in Python
Excellent communication skills
Preferred Qualifications
PhD in Data Science, Statistics, or a quantitative social science field
Hands-on industry experience with product-focused statistical analysis
Experience working with large-scale multimodal data and data-annotation pipelines
Experience with LLM prompt engineering & prompt optimization
Experience with LLM auto-judges for generative AI model evaluation
A track record of publications or technical presentations in Data Science or a related field
Excellent at cross-functional collaboration
Skills
Machine Learning
Generative AI
LLM
Python
Similar jobs
Data Scientist
Leidos · McLean, United States
27 minutes ago$154.1k - $278.5k/yrStaff Data Scientist, Marketing
Asana · United States
33 minutes ago$202k - $282k/yrLead Data Scientist - Growth & Marketing Models
FairSquare · United States
33 minutes ago$150k - $170k/yrData Scientist
The Coca-Cola Company · United States
1 hour ago$109k - $129k/yrSAP Data Scientist
Accenture LLP · Washington, United States
1 hour ago$116.9k - $243.1k/yrAIML - Sr Data Scientist, Evaluation
Apple, Inc. · Cupertino, United States
1 hour ago