Why This Role Stands Out
As a Staff AI/Machine Learning Engineer at Tonicai, you'll design and build critical systems that power the future of AI, working on complex, real-world data challenges with leading AI labs and enterprises. This remote role offers immense growth potential and the chance to develop cutting-edge skills in synthetic data generation and de-identification, making it ideal for ambitious engineers who thrive on impactful innovation. Apply today to shape the data infrastructure behind modern AI!
Quick Overview
Job Description
About Tonic
Tonic builds the data infrastructure behind modern AI. We generate the synthetic environments that agents are trained and tested in, and we de-identify real enterprise data so it can be used safely in training and evaluation. Eight years in, we work with frontier AI labs pushing the edge of what models can do, and with hundreds of enterprises including Fidelity, Comcast, eBay, and Vanguard, on the data problems that sit at the center of where AI is going next.
About The Role
The models you build here are load-bearing. The environments you generate decide whether an agent is ready to ship or only looked good in a demo. The synthesis and de-identification models you train decide whether a bank can safely put its data near a model at all. And the work spans real range: in one week you might build evaluation that separates the best models from the rest on real tasks, train a synthesis model where both fidelity and downstream utility have to hold, and improve entity detection on messy production data. Real enterprise data, real stakes, and problems that don’t have textbook answers yet.
What You'll Do
Design and build the systems that generate longitudinally coherent synthetic environments for agent training and evaluation, including persona modeling, task generators, and verifiable ground truth.
Build and maintain synthesis models that generate realistic replacement values at very large scale, preserving format, statistical distribution, and semantic consistency so de-identified data stays useful downstream.
Train and improve the NER models behind our entity detection, driving accuracy and recall across free text, structured fields, and mixed enterprise data at scale.
Build evaluation infrastructure that grades agent outcomes, not just traces, and produces real discrimination between frontier models on real tasks.
Fine-tune and evaluate open-weight models on Tonic-generated data, and turn benchmark results into product and research direction.
Expand coverage into new domains, languages, and entity types, and handle the long tail of formats and edge cases that real customer data throws off.
Own model evaluation across the board: precision and recall on detection, utility preservation on synthesis, and outcome-level grading for agents.
Optimize inference so models run efficiently on large volumes of sensitive data inside customer environments.
Partner directly with frontier labs and enterprise ML team to turn hard data problems into shipped model improvements.
What You’ll Bring
8+ years (or PhD with 3+ years) building production ML systems, with real depth in some combination of LLMs, agents, RL, NER, or information extraction.
Hands-on experience training and shipping models to production, and a pragmatic bar for quality: you know how to measure it, where it breaks, and when it's good enough to ship.
Experience with generative or synthesis models where output fidelity and downstream utility both matter, not just plausibility.
Strong software engineering fundamentals. You write code others build on.
Fluency with modern training and eval stacks (PyTorch, distributed training, standard agent and benchmark frameworks).
Comfort working with messy, sensitive, real-world data and the privacy constraints that come with it.
A track record of framing ambiguous problems and driving them to measurable, shipped results.
Bonus: synthetic data generation, data privacy or de-identification, or benchmark construction.
Benefits We Offer
Competitive salary and equity
Unlimited paid time off
401k plan with employer contribution
Medical, dental, and vision insurance
Generous parental leave policy
Remote-friendly work environment
Similar jobs
- NA
Senior/Lead Bioinformatics Scientist (Multi-Cancer Early Detection)
NewAuto ApplyNatera
US Remote🇺🇸RemoteYesterdayPython - NA
Staff Machine Learning and Bioinformatics Scientist (Multi-Cancer Early Detection)
NewAuto ApplyNatera
US Remote🇺🇸RemoteYesterdayMachine LearningDeep LearningPythonTechnology - NA
Machine Learning Scientist, Multimodal AI
NewAuto ApplyNatera
US Remote🇺🇸$124.8k - $171.6k/yrRemoteYesterdayAWSMachine LearningDeep Learning+2Technology - CL
Machine Learning Engineer (Mid-Level)
NewAuto ApplyClera
San Francisco🇺🇸On-site22 hours agoDockerGCPAWS+8Technology - FA
Applied AI/ML Scientist, Intern
NewAuto ApplyFaire
San Francisco🇺🇸$75/hrHybridYesterdaySQLMachine LearningDeep Learning+3Technology - WH
Staff Applied Machine Learning Scientist (Health)
NewAuto ApplyWhoop
Boston🇺🇸On-siteYesterdayMachine LearningNumPySciPy+6Technology - WH
Senior Manager, ML Health - Insights
NewAuto ApplyWhoop
Boston🇺🇸$170k - $230k/yrOn-siteYesterdayMachine LearningOnboardingPerformance Management - CO
Staff Machine Learning Engineer
NewAuto ApplyCoinbase
Remote - USA🇺🇸$218.0k/yrRemoteYesterdayMachine LearningNLPArbitration+4Technology - CA
Senior Machine Learning Engineer
NewAuto ApplyCare Access
USA Remote🇺🇸$140k - $190k/yrRemoteYesterdayDockerSQLAWS+16Technology - AL
Senior Scientist, Machine Learning
NewAuto ApplyAltos Labs
San Francisco Bay Area🇺🇸Hybrid23 hours agoExpressMachine LearningDeep Learning+4Technology - AI
Staff MLOps Engineer
NewAuto ApplyAiDASH, Inc.
Palo Alto🇺🇸Hybrid22 hours agoDockerAWSMLOps+1Engineering - PE
Applied Machine Learning Engineer
NewAuto ApplyPermitflow
New York City🇺🇸HybridYesterdayGCPAWSMachine Learning+10Technology