4 hours ago
Base Salary
$165k - $270k/yr
Responsibilities
- Manage GPU and CPU infrastructure deployments to Top Secret data centers.
- Provide GPU-as-a-service support for external customers on bare-metal and virtualized platforms.
- Design, validate, and productize AI cluster solutions at 100,000+ GPU scale.
- Develop automation for deploying and managing on-premise Kubernetes and AI clusters and operating systems.
- Deploy and manage databases, monitoring systems, and distributed storage.
- Collaborate with AI engineers to build scalable, operable, and maintainable products.
- Manage the full service lifecycle from design and deployment through operation and refinement.
- Improve monitoring, alerting, and system availability.
- Identify infrastructure improvements and create innovative high-availability solutions.
- Mentor junior engineers and lead the team toward technical excellence.
Requirements
- Bachelor’s degree in computer science, information systems/IT, or engineering plus 5+ years of professional Linux operating-system experience, or 7+ years of software, DevOps, or site reliability engineering experience in lieu of a degree.
- At least 5 years of Kubernetes experience and 5 years of Linux operating-system management experience.
- Experience with Terraform, Ansible, or comparable infrastructure tools and containerization technologies.
- Scripting experience in Bash, Python, or similar languages and development experience in Python, C++, or Go.
- Preferred qualifications include 5+ years of Python and Python-based development, Kubernetes cluster management, Linux boot and systems configuration knowledge, and experience managing thousands of servers.
- Preferred qualifications include knowledge of testing, continuous integration, build and deployment technologies, distributed databases, data modeling, TCP/IP networking, cloud virtualization, and NVIDIA GPU deployment stacks.
- Active Top Secret, Top Secret SCI, or DOE Level Q clearance is preferred.
- Must be willing to work extended hours and weekends, travel domestically and globally as needed, and successfully obtain and maintain a Top Secret security clearance.
- Applicants must meet applicable ITAR eligibility requirements or be eligible to obtain required authorizations.
Benefits
- Base salary range is $165,000-$270,000 annually, with potential stock or long-term cash awards, discretionary bonuses, and an Employee Stock Purchase Plan.
- Benefits include medical, vision, dental, 401(k), disability and life insurance, paid parental leave, approximately three weeks of paid vacation, 10 or more paid holidays, and applicable paid sick time.
- Employees with an active clearance may receive a 10% differential up to an additional $20,000 annually after being briefed into a classified program.
- Company shuttles operate Monday through Friday from select Seattle locations to the SpaceX Redmond office.
- The role may require extended hours, weekends, and future domestic or global travel.
About SpaceX
SpaceX designs, manufactures, and launches orbital rockets and spacecraft, and operates Starlink, a global satellite internet network for consumers, businesses, and governments. Its revenue comes from commercial and government launch services (Falcon 9/Falcon Heavy, Dragon cargo and crew to the ISS) and subscription broadband with Starlink hardware and service. Founded in 2002 and headquartered in Hawthorne, California, the privately held company serves NASA and commercial satellite operators, building most systems in-house.
