Why This Role Stands Out
As a Member of Technical Staff at Liquid AI, you'll contribute to the core inference systems that power cutting-edge AI, offering significant growth and impact within a rapidly scaling, MIT-spun-out company. This remote hybrid role is ideal for ambitious mid-senior engineers who excel at quickly mastering new technologies and possess a keen eye for performance optimization and rigorous benchmarking. You'll have the opportunity to collaborate directly with research, product, and external engineering teams, shaping the future of AI deployment.
Quick Overview
Job Description
Liquid AI Job Description
Role: Member Of Technical Staff, Infrastructure
Department: Research & Engineering
Location: Boston
Location Type: Hybrid
Employment Type: Full-time
About Liquid AI
Spun out of MIT CSAIL, we build general-purpose AI systems that run efficiently across deployment targets, from data center accelerators to on-device hardware, ensuring low latency, minimal memory usage, privacy, and reliability. We partner with enterprises across consumer electronics, automotive, life sciences, and financial services. We are scaling rapidly and need exceptional people to help us get there.
The Opportunity
Our inference stack is central to everything we ship. You'll be a core part of the team responsible for the engine layer that runs our models in production and in partner environments, and for the benchmarking infrastructure we use to evaluate our own work and verify what partners bring to us. Day to day, that means working closely with research and product, but also directly with external engineering teams.
What We're Looking For
We need someone who:
Can pick up unfamiliar tools quickly and knows how to assess whether they're worth using.
Designs AI benchmarks and holds methodology to a high standard.
Cares about inference details, understands the tradeoffs, and checks what changed across the board before calling something done.
Doesn’t consider a model port finished until you can prove the outputs are correct.
The Work
Design and build benchmark suites that cover inference performance, model quality, and knowledge evaluation across different hardware targets.
Run external partner verifications: evaluate their solutions against our benchmarks, identify gaps, and clearly deliver findings.
Port models like LFM2 onto different runtimes and frameworks, and verify correctness end-to-end.
Maintain and extend the inference engine layer built on llama.cpp, ONNX, and MLX as new model architectures emerge from research.
Make benchmark results explainable and verifiable, so internal teams and partners can trust and reproduce them independently.
Desired Experience
Must-have:
Hands-on experience with at least one inference framework like llama.cpp, ONNX Runtime, or MLX, going beyond basic usage into internals and modification.
Experience designing and building benchmarking pipelines, including methodology, validation, and reproducibility.
Strong C++ and Python in performance-sensitive contexts.
Solid understanding of inference fundamentals: quantization, decoding strategies, memory layout, and how they interact.
Nice-to-have:
Experience porting models across runtimes and verifying numerical correctness.
Prior work with external partners or clients in a technical validation or evaluation capacity.
Familiarity with edge inference targets and the constraints that come with them.
What Success Looks Like (Year One)
You've ported LFM2 onto multiple runtimes and platforms, you know the model inside out, and new ports take you a fraction of the time they did at the start.
You've run multiple partner verifications end-to-end and built enough context to spot weak evaluations quickly and push back with evidence.
The benchmark suite covers inference performance and model quality across the platforms we care about, and both internal teams and partners are using it as a reference.
What We Offer
Compensation: Competitive base salary with equity in a unicorn-stage company
Health: We pay 100% of medical, dental, and vision premiums for employees and dependents
Financial: 401(k) matching up to 4% of base pay
Time Off: Unlimited PTO plus company-wide Refill Days throughout the year
Similar jobs
- ZT
Software Engineer III - Developer Productivity
NewZoomInfo Technologies LLC
Bethesda🇺🇸3 hours agoDockerGCPMicroservices+13Technology - ZT
Software Engineer III
ZoomInfo Technologies LLC
Bethesda🇺🇸2 weeks agoGCPSpringAWS+2Technology - VE
Software Engineer - Network and Automation
NewVerisign
Reston🇺🇸$135.8k - $183.8k/yr9 hours agoDockerAnsibleDNS+5Technology - SK
Principal Software Engineer (Android)
NewStable Kernel
Atlanta🇺🇸6 hours agoMongoDBMySQLAWS+16Technology - SC
Senior Software Engineer
Schonfeld
New York🇺🇸7 weeks agoDockerMicroservicesAWS+11Technology - RE
Senior Software Engineer, Consumer Engineering
NewReddit
Remote - United States🇺🇸$190k - $267k/yrRemote4 hours agoFirebaseRustAccount Management+15Technology