2 days ago
San Francisco, CA, USAMid Level
Base Salary
$150k - $200k/yr
Responsibilities
- Build and iterate on the cloud-based ASR pipeline from audio capture through post-processing in production at scale.
- Own ASR quality and reliability, including improvements to latency, small-word accuracy, and voice-print reliability.
- Work across data preparation, model training and fine-tuning, evaluation, and deployment to ship pipeline improvements.
- Debug transcription quality issues in live systems using production usage data.
- Collaborate with product, overseas R&D, hardware, and supply-chain teams across time zones.
- Turn lightly specified product needs into concrete, production-ready improvements.
Requirements
- At least 3 years of experience building and tuning end-to-end transcription or ASR pipelines in production, primarily in cloud environments.
- Demonstrated ownership of production ASR systems across data preparation, model training or fine-tuning, evaluation, and deployment.
- Hands-on experience with latency-sensitive or streaming audio and ASR pipelines.
- Experience shipping pipeline improvements from design through deployment based on real production usage data.
- Experience debugging transcription quality, latency, and voice-print reliability issues in live systems.
- Experience in early-stage or founding engineering environments with minimal team support and specification.
- Experience with on-device or embedded ML, such as Core ML or TensorFlow Lite.
- Prior experience building wearable, hardware, or robotics device products.
- Background at an AI-native consumer application focused on transcription or audio.
- Experience building agent or LLM-based product features involving tool use, memory, or retrieval systems.
Benefits
- Base salary of $150,000–$200,000 USD annually.
- Hybrid schedule with 3 days per week in the San Francisco Bay Area office.
