Why This Role Stands Out
As an ASR Engineer at Clera, you'll take end-to-end ownership of a critical cloud-based pipeline, shaping the future of ambient intelligence and enjoying significant impact in one of the company's first US engineering roles. This is a fantastic opportunity for a mid-senior engineer with production ASR experience to drive innovation and make tangible improvements in a dynamic, early-stage startup environment. Apply now to build and iterate on cutting-edge technology!
Quick Overview
Job Description
About the Role
This is an end-to-end ownership role for a cloud-based ASR and transcription pipeline at an early-stage ambient intelligence consumer startup. You'll work directly with product and general management leadership as one of the company's first US engineering hires, making real tradeoffs between latency, accuracy, and reliability as the product evolves.
What You'll Do
Build and iterate on the cloud-based ASR pipeline, from audio capture through post-processing, in production at scale.
Own ASR quality and reliability end-to-end, shipping measurable improvements on latency, small-word accuracy, and voice-print reliability.
Work across data, training and fine-tuning, evaluation, and deployment to turn product feedback into shipped pipeline changes.
Collaborate closely with overseas R&D, hardware, and supply-chain teams across time zones.
Partner with a product engineer on shared backend and pipeline surfaces.
Operate with minimal specification, translating lightweight asks into concrete, production-ready improvements.
What We're Looking For
3+ years building and tuning transcription and ASR pipelines end-to-end in production, primarily in cloud-based settings.
Demonstrated ownership of production ASR systems through the full lifecycle: data preparation, model training and fine-tuning, evaluation, and deployment.
Hands-on experience with latency-sensitive or streaming audio and ASR pipelines.
Proficiency across the ML lifecycle, including data handling, evaluation metrics, and production deployment.
Track record of debugging and tuning transcription quality issues such as small-word accuracy, voice-print reliability, and latency.
Experience in early-stage or founding engineering environments, shipping without large team support or fully-specified requirements.
On-device or embedded ML experience (Core ML, TensorFlow Lite, or similar frameworks) is a plus.
Prior experience with wearable, hardware, or robotics products is a plus.
Background at AI-native consumer applications focused on transcription or audio is a plus.
Experience building agent or LLM-based product features, including tool use, memory, or retrieval, is a plus.
You care about how transcription feels to use, not just how it benchmarks, and can make latency and accuracy tradeoffs independently.
Compensation & Benefits
Salary range: $150,000 to $200,000 USD annually. Visa sponsorship is not available for this role.
Location
Hybrid, 3 days per week in office. San Francisco Bay Area, California, United States.
Similar jobs
- JM
Lead Software Engineer
NewJ.P. Morgan
Columbus, Ohio🇺🇸On-site10 minutes agoDockerFastAPIFlask+15Technology - BO
Senior Systems Engineer (Reliab, Maintain & Sys Health)
NewBoeing
Saint Louis, Missouri🇺🇸$164.9k - $223.1k/yrHybrid50 minutes agoTechnology - VA
Avionics System Test Engineer
NewVast
Long Beach🇺🇸6 hours agoSpringEmbedded SystemsCompliance+8Technology - VE
Senior Software Engineer - Camera Platform
NewVerkada
San Mateo🇺🇸9 hours agoEdge ComputingBashC+++5Technology - MA
Application Developer
NewMANTECH
Chantilly, Virginia🇺🇸Hybrid55 minutes agoSQLSpringAngular+6Technology - JM
Lead Software Engineer - Backend Lead
NewJ.P. Morgan
Plano, Texas🇺🇸On-site55 minutes agoGCPMicroservicesOracle+13Technology