1 month ago
Base Salary
$150k - $215k/yr
Responsibilities
- Design, develop, and operate Kubernetes clusters running on AWS infrastructure.
- Improve the scalability, resiliency, performance, and security of the AI agent runtime environment.
- Build systems and tooling for running long-running ad hoc tasks on cost-efficient infrastructure.
- Automate portions of the software development lifecycle and lead developer productivity improvements through tooling, automation, and platform integrations.
- Partner with application, security, and data teams to implement secure-by-default infrastructure practices.
- Participate in incident response, root cause analysis, and postmortems to improve platform reliability.
- Mentor engineers, provide technical leadership across the Platform organization, and help define the AI DevEx team roadmap.
Requirements
- At least 5 years of experience in DevOps, platform, site reliability, cloud engineering, or backend software engineering roles.
- Deep understanding of Kubernetes architecture and core components.
- Knowledge of AI runtimes and their environmental requirements.
- Hands-on experience operating cloud infrastructure, preferably AWS.
- Hands-on experience with Infrastructure as Code tools such as Terraform.
- Ability to evaluate system performance, identify bottlenecks, and use data to drive improvements.
- Experience collaborating with multiple stakeholders and prioritizing work for business impact.
Benefits
- The role is based in WHOOP’s Boston, Massachusetts office, and the successful candidate must be prepared to relocate if necessary.
- WHOOP is an Equal Opportunity Employer and participates in E-Verify.
Tech Stack
Categories
About WHOOP
Solving problems that sit at a unique intersection of hardware, tech, health, fitness and design.
