Relace

Infrastructure Engineer

Relace
Apply
11 months ago

Responsibilities

  • Architect and manage infrastructure for the inference and training stack.
  • Build reliable and efficient systems for deploying and scaling machine learning workloads globally.
  • Work on GPU scheduling, distributed systems, and high-performance cloud deployments.
  • Optimize performance and cost across compute, networking, and storage layers.
  • Collaborate with research and product teams to operate models at scale.

Requirements

  • At least 2 years of experience writing high-quality production code.
  • Strong experience with cloud infrastructure such as AWS, GCP, Azure, or equivalent.
  • Experience with data science and systems optimization.
  • Familiarity with machine learning infrastructure and GPUs is a plus.

Benefits

  • Work out of the company's San Francisco office in the Financial District.
Relace

About Relace

1-10 employees
Contact me