7 months ago
San Francisco, CA, USAMid Level
Responsibilities
- Design, implement, and iterate on models across the voice stack, including audio, speech, and language.
- Build training pipelines, evaluation frameworks, and tooling for rapid experimentation.
- Translate research results into production-grade systems with the engineering team.
- Form hypotheses, investigate results, and take ownership of assigned problems.
- Collaborate with researchers and product engineers to advance Phonic’s voice AI systems.
Requirements
- Hands-on experience in machine learning research or research engineering in industry or PhD-level academia.
- Proficiency in Python and PyTorch or JAX, with the ability to implement models from research papers.
- Experience running machine learning experiments end to end, including data processing, training, evaluation, and iteration.
- Ability to work in fast-moving, ambiguous environments where defining the problem is part of the role.
- Ability to clearly explain experimental approaches, results, and reasoning.
- Research experience in speech, audio, or language modeling such as ASR, TTS, LLMs, or codec models is preferred.
- Familiarity with diffusion, flow matching, or autoregressive generative models is preferred.
- Experience with distributed training, quantization, or inference optimization is preferred.
- Competitive programming or olympiad experience is preferred.
Benefits
- Fully in-person work in the San Francisco office
- Free breakfast, lunch, and dinner provided in the office
- Comprehensive health, dental, and vision coverage
- Regular off-sites and team celebrations
- 401(k) plan
Categories
AI ResearchML Engineering
