11 months ago
San Francisco, CA, USAMid Level
Responsibilities
- Architect and manage infrastructure for the inference and training stack.
- Build reliable and efficient systems for deploying and scaling machine learning workloads globally.
- Work on GPU scheduling, distributed systems, and high-performance cloud deployments.
- Optimize performance and cost across compute, networking, and storage layers.
- Collaborate with research and product teams to operate models at scale.
Requirements
- At least 2 years of experience writing high-quality production code.
- Strong experience with cloud infrastructure such as AWS, GCP, Azure, or equivalent.
- Experience with data science and systems optimization.
- Familiarity with machine learning infrastructure and GPUs is a plus.
Benefits
- Work out of the company's San Francisco office in the Financial District.
