- Clera
- San Francisco, CA
- Full-Time
- 23 days ago
ASR Engineer.
Before you go
Before you leave us, sign up for our email alerts
We don't do job spam, just the best digital jobs delivered straight to your inbox.
ASR Engineer: our view in 3 lines...
- The Role:This role is for an engineer building and improving a cloud-based ASR system with a small on-device component for an early-stage ambient intelligence consumer startup.
- The Person:The person will build, tune, debug, and ship improvements to the transcription pipeline, owning quality and reliability across data, training, evaluation, deployment, and production debugging.
- Requirements:The role calls for 3+ years building and tuning transcription or ASR pipelines in production, hands-on experience with latency-sensitive or streaming audio pipelines, on-device or embedded ML experience, and experience building agent or LLM-based product features.
About the role
About the Role
You'll own the transcription pipeline end-to-end at an early-stage ambient intelligence consumer startup — a cloud-based ASR system with a narrowly scoped on-device component. As one of the company's first US engineering hires, you'll work hands-on with senior product and business leadership to build, tune, debug, and ship pipeline improvements directly.
What You'll Do
-
Build and iterate on the cloud-based ASR pipeline, from audio capture through post-processing, in production at scale.
-
Own ASR quality and reliability end-to-end, shipping measurable improvements across latency, small-word accuracy, and voice-print reliability.
-
Work across data, training/fine-tuning, evaluation, and deployment to translate product feedback into shipped pipeline changes.
-
Collaborate closely with overseas R&D, hardware, and supply-chain teams across time zones.
-
Partner with a product-focused backend engineer on shared pipeline surfaces.
-
Operate with minimal specification, turning lightweight asks into concrete, production-ready improvements.
What We're Looking For
-
3+ years building and tuning transcription/ASR pipelines end-to-end in production, primarily in cloud-based settings.
-
Demonstrated ownership of production ASR systems through the full lifecycle: data preparation, model training/fine-tuning, evaluation, and deployment.
-
Hands-on experience with latency-sensitive or streaming audio/ASR pipelines.
-
Track record of shipping pipeline improvements from design through deployment based on real production usage data.
-
Comfort debugging transcription quality issues (small-word accuracy, voice-print reliability, latency) in live systems.
-
Experience in early-stage or founding engineering environments — ships without large team support or fully-specified requirements.
-
On-device or embedded ML experience (Core ML, TensorFlow Lite, or similar frameworks).
-
Prior experience building wearable, hardware, or robotics device products.
-
Background at an AI-native consumer application focused on transcription or audio.
-
Experience building agent or LLM-based product features (tool use, memory, retrieval systems).
-
Cares how transcription feels to use — not just how it benchmarks — and makes latency/accuracy tradeoffs independently.
Compensation & Benefits
Base salary $150,000 – $200,000 USD annually. Visa sponsorship is not available for this role.
Location
Hybrid (3 days/week in office) — San Francisco Bay Area, California, USA.

