5 months ago
Responsibilities
- Optimize machine learning models through pruning, quantization, and distillation for edge and cloud deployment.
- Deploy models through high-performance APIs and integrate them with the broader software ecosystem.
- Build and maintain machine learning inference frameworks that support AI research at scale.
- Collaborate with AI Scientists to refactor research code into production-grade, maintainable software.
- Focus on production latency, reliability, data integrity, and hardware-aware optimization.
Requirements
- 5+ years of software engineering experience or equivalent university time with deep experience in ML-specific systems and distributed computing.
- Proficiency in C++ and Python.
- Experience with Docker and Kubernetes.
- Experience with vector databases including Pinecone and Weaviate.
- Bachelor’s or master’s degree in Computer Science or Software Engineering.
Benefits
- Competitive salary.
- Comprehensive health, dental, and vision benefits for U.S.-based employees, with benefits varying internationally based on local requirements and employing entity.
- 401(k) match and equity options.
- $200/month Health & Wellness stipend.
- Continuing Education support and a $500/year Function Health subscription.
- Free parking for in-office employees.
- Flexible Time Off, parental leave for eligible employees, and supplemental life insurance.
