5 months ago
Base Salary
$150k - $450k/yr
Responsibilities
- Design and execute large-scale speech data curation and processing pipelines for real-world audio, synthetic data, and automated annotation.
- Perform pre-training and post-training of speech-language models to improve accuracy, factuality, natural spoken style, conversational tone, and multilingual fluency.
- Build evaluation frameworks covering quality, latency, expressiveness, factuality, human preferences, real-time interaction, and experimentation.
- Integrate voice models into applications and real-time environments, defining spoken interaction specifications and managing the lifecycle from prototype through global-scale deployment.
Requirements
- Expert proficiency in Python for clean and efficient AI/ML systems code.
- Hands-on experience processing large-scale datasets with tools such as Spark and Ray.
- Experience pre-training and post-training speech-language models using JAX or PyTorch, including supervised fine-tuning and reinforcement learning.
- Ability to build and run rigorous evaluation pipelines covering objective metrics, human preference studies, factuality checks, and A/B testing.
- Experience with large-scale distributed training and inference systems on Kubernetes.
- Proactive, self-driven execution in a fast-paced, high-caliber team.
Benefits
- Equity
- Comprehensive medical, vision, and dental coverage
- 401(k) retirement plan
- Short- and long-term disability insurance
- Life insurance
- Various discounts and perks
Tech Stack
Categories
AI ResearchML Engineering
About xAI
Understand the Universe. We are a team of AI technologists and business leaders on a mission to build AI systems that can help humanity understand the world better. https://x.ai/careers