Overview
ABOUT THE ROLE
Full job description
ABOUT THE ROLE This is a foundational engineering role at an early-stage AI consumer hardware and software startup, where you will own the transcription pipeline end-to-end. You will work hands-on with product and general management leadership to build, tune, and ship a cloud-based ASR system with a narrowly scoped on-device component. Your work directly shapes how well the core product experience feels to real users. WHAT YOU'LL DO
- Build and iterate on the cloud-based ASR pipeline, from audio capture through post-processing, running in production at scale.
- Own ASR quality and reliability end-to-end, shipping measurable improvements across latency, small-word accuracy, and voice-print reliability.
- Work across data preparation, model training and fine-tuning, evaluation, and deployment to translate product feedback into shipped pipeline changes.
- Collaborate with a Partner Product Engineer on shared backend and pipeline surfaces.
- Coordinate across time zones with R&D, hardware, and supply-chain teams based in China.
- Operate with minimal specification, turning informal asks into concrete, shipped improvements. WHAT WE'RE LOOKING FOR
- 3 or more years building and tuning transcription and ASR pipelines end-to-end in production, primarily in cloud-based settings.
- Demonstrated ownership of production ASR systems across the full lifecycle: data preparation, model training and fine-tuning, evaluation, and deployment.
- Experience building and optimizing latency-sensitive or streaming audio and ASR pipelines.
- Track record of making latency, accuracy, and reliability tradeoffs based on real user feedback.
- Experience debugging and tuning transcription quality issues in production environments.
- Comfort shipping in early-stage or founding engineering environments with small teams and limited specification.
- On-device or embedded ML experience using frameworks such as Core ML or TensorFlow Lite.
- Prior experience with wearable, hardware, or robotics device products.
- Background at AI-native consumer applications focused on transcription or audio.
- Experience building agent or LLM-based product features including tool use, memory, or retrieval systems.
- Ability to work hybrid three days per week in the San Francisco Bay Area.
- Ability to collaborate asynchronously with international teams across time zones. COMPENSATION & BENEFITS Salary range: $150,000 to $200,000 USD annually. Visa sponsorship is not available for this role. LOCATION Hybrid, three days per week on-site in the San Francisco Bay Area, California, United States.
Tips for this job
Practical JobOpportunity guidance. These tips do not replace official rules or create new eligibility requirements.
- Tailor the CV and application to the responsibilities and required skills stated on the official employer page.
- Use concrete evidence of relevant work, projects and measurable results rather than generic claims.
- Confirm location, work authorization, remote restrictions and sponsorship terms before applying.
- Apply through the original employer or official recruitment destination shown on this page.
Verification notes
laptop-ats-crawler v3
JobOpportunity is the discovery and verification layer. Confirm eligibility, dates, salary/funding and application instructions on the original source before submitting anything.
Apply through JobOpportunity →