
Member of Technical Staff - AI Cloud Infrastructure
Emerald AIabout 2 hours ago
Boston, MA, USA +2 moreSenior / Staff+
Responsibilities
- Architect managed services from 0→1, defining productization of GPU capacity.
- Engineer the platform core, building control-plane services and customer interfaces.
- Onboard and vet infrastructure partners through technical assessments.
- Design end-to-end multi-tenancy with rigorous isolation across compute, storage, and networking.
- Drive workload orchestration in Kubernetes and Slurm environments.
- Lead high-performance storage strategy by deploying parallel storage solutions.
- Ensure operational excellence by defining SLOs and incident response protocols.
Requirements
- At least 7+ years of experience in infrastructure or platform engineering.
- Strong experience with Kubernetes and Slurm as managed services.
- Production experience with Lustre or comparable parallel filesystems.
- Solid understanding of cloud service fundamentals and operational discipline.
- Deep Linux systems knowledge and experience with infrastructure as code tools.
- Familiarity with GPU infrastructure and high-performance networking.
Benefits
- Make an impact by solving the AI power bottleneck.
- Join a collaborative, low-ego team of experts.
- Influence strategy and customer engagement from day one.
- Competitive pay and equity options.
- Comprehensive benefits including medical, dental, vision, and 401(k) matching.
- Flexible work location with 2 WFH days per week.