
Software Engineer: ML Infra
Generalist7 months ago
Base Salary
$200k - $350k/yr
Responsibilities
- Own the company’s GPU compute fleets.
- Make GPUs easy for researchers to use and maximize their utilization.
- Optimize ML data loading, transport, and storage in highly distributed, fully utilized environments.
- Orchestrate inference fleets running on on-premise GPUs attached to robots.
- Operate infrastructure for distributed training jobs, researcher experiments, and latency-sensitive robot inference.
Requirements
- Experience managing large GPU fleets for large-scale, long-term, highly distributed training runs or inference.
- Deep experience with Slurm or Kubernetes for ML workload orchestration.
- Experience building high-scale ML data loaders and data preparation systems.
- Deep understanding of ML hardware, storage, and networking stacks.
- Experience in the NVIDIA GPU ecosystem.
Tech Stack
Categories
About Generalist
Generalist builds general‑purpose robots powered by large‑scale AI, focusing on vision‑language‑action and embodied multimodal models for real‑world tasks in industry and homes. The company develops robotics platforms and software and is privately held. It was founded in 2024, is headquartered in San Mateo, California, and its founding team includes alumni of OpenAI, Google DeepMind, and Boston Dynamics.