2 months ago
Remote, United States or Austin, TX, USAStaff+
Responsibilities
- Lead development and optimization of Large Language Models and Mixture of Experts models.
- Lead initiatives involving on-device inference optimization, quantization, RAG, agentic frameworks, and tool calling.
- Integrate machine learning models into the platform with cross-functional engineering teams.
- Conduct research focused on improving the performance and efficiency of LLMs.
- Lead the sub-team and deliver initiatives end to end.
- Mentor junior engineers and contribute to knowledge sharing and best practices.
Requirements
- Advanced degree in Computer Science or a related field; a Ph.D. is preferred.
- At least 6 years of machine learning experience with specific expertise in Large Language Models and Mixture of Experts.
- Proven track record of building and innovating through publications or industry experience.
- Strong programming skills in Python and experience with TensorFlow and/or PyTorch.
- Ability to lead complex projects and work collaboratively in a team environment.
- Strong problem-solving skills and a passion for innovation.
- Preferred: experience with AWS, Azure, or GCP.
- Preferred: knowledge of Hadoop or Spark.
- Preferred: familiarity with Docker or Kubernetes.
- Preferred: publications or presentations in recognized machine learning journals or conferences.
Benefits
- Competitive benefits package including health, dental, and vision coverage.
- 401(k) match and equity options.
- Health and wellness stipend.
- Continuing education support.
- Function Health subscription.
- Free parking for in-office employees.
- Flexible Time Off.
- Parental leave for eligible employees.
- Supplemental life insurance; benefits may vary for employees hired outside the United States.
Tech Stack
Categories
AI ResearchML Engineering
