5 months ago
Base Salary
$205k - $270k/yr
Responsibilities
- Design, train, evaluate, and deploy machine learning systems for ASR, speech understanding, turn detection, text to speech, speech to speech, classification, entity extraction, summarization, and structured insight generation.
- Improve voice AI quality through error analysis, data curation, metric design, benchmarking, and iterative model improvement.
- Build evaluation frameworks for voice and agentic systems using metrics including accuracy, robustness, latency, faithfulness, naturalness, professionalism, task completion, and cost.
- Diagnose and mitigate transcription errors, hallucinations, retrieval failures, tool misuse, prompt brittleness, context drift, and multi-step reasoning failures.
- Design and optimize low-latency ML workflows for live conversations while balancing quality, responsiveness, scalability, and reliability.
- Partner with platform and backend engineers to productionize real-time inference, streaming pipelines, quality monitoring, and continuous model iteration.
- Collaborate with product, design, frontend, and backend teams to integrate voice intelligence into Cresta’s platform.
- Establish best practices for offline evaluation, online experimentation, model validation, observability, and production quality monitoring.
- Mentor engineers, contribute to technical strategy, and help shape the roadmap for Cresta’s voice AI systems.
Requirements
- Bachelor’s degree in Computer Science, Mathematics, Machine Learning, AI, or a related field is required; a Master’s or Ph.D. is preferred.
- At least 5 years of experience building, evaluating, and deploying machine learning systems in production.
- Strong background in speech recognition, speech processing, NLP, generative AI, or conversational AI.
- Deep experience with model evaluation, benchmarking, error analysis, and production ML quality improvement.
- Expertise with modern ML frameworks and tooling such as PyTorch, TensorFlow, and Hugging Face.
- Understanding of transformer-based models, embeddings, retrieval systems, and large-scale training or inference workflows.
- Experience designing and deploying real-time ML systems with latency, scalability, and reliability requirements.
- Experience building data pipelines and tooling for experimentation, measurement, and large-scale quality analysis.
- Experience translating research ideas into production-grade systems and working across research and engineering boundaries.
- Strong communication and technical leadership skills.
- Experience with ASR quality metrics such as WER and task-level evaluation is preferred.
- Experience with RAG systems, agentic workflows, multi-step reasoning systems, or LLM-as-a-judge evaluation is preferred.
- Familiarity with streaming inference, real-time voice pipelines, or media systems is preferred.
- Experience partnering with infrastructure or platform teams on ML deployment, observability, and reliability is preferred.
- Experience in contact center AI, conversational intelligence, or enterprise voice products is preferred.
Benefits
- Comprehensive medical, dental, and vision coverage
- Flexible paid time off
- Paid parental leave for all new parents
- Retirement savings plan
- Remote work setup budget
- Monthly wellness and communication stipend
- In-office meal program and commuter benefits for onsite employees
Tech Stack
PyTorchTensorFlow
Categories
About Cresta
Cresta builds a generative AI platform for contact centers, combining AI agents, real-time agent assistance, conversation intelligence, and quality/coaching tools. It sells enterprise software to customer-support and sales organizations to automate and improve conversations across voice and digital channels. Founded in 2017 and headquartered in Sunnyvale, California, the privately held company counts brands such as Alaska Airlines, Cox Communications, Intuit, United Airlines, and Marriott among its customers.
