10 hours ago
Responsibilities
- Build and enhance compute infrastructure for Apple’s global services.
- Design and operate VM lifecycle management and bare-metal infrastructure on Apple’s Cloud Platform.
- Build and scale cloud infrastructure and large-scale distributed systems.
- Develop automation and AI-assisted tools for triage, operational workflows, and SRE infrastructure operations.
- Manage, scale, and troubleshoot highly distributed and potentially planet-scale, multi-tenant infrastructure.
- Collaborate with multiple engineering teams and mentor other engineers.
Requirements
- Bachelor’s degree in Computer Science, an engineering-related field, or equivalent related experience.
- At least 7 years of experience in an infrastructure-focused Site Reliability Engineering role.
- Strong experience building and scaling cloud infrastructure and large-scale distributed systems.
- Experience with OpenStack, KVM or hypervisor technologies, and Kubernetes.
- High proficiency in Go or Python.
- Experience with highly distributed Unix systems.
- Strong verbal and written communication, ownership, and automation skills.
- Preferred: advanced degree, planet-scale operations experience, large-scale multi-tenant Infrastructure as a Managed service experience, and expert proficiency with Chef, Ansible, or Terraform.
Tech Stack
Categories
DevOpsSite Reliability
About Apple
We’re a diverse collective of thinkers and doers, continually reimagining what’s possible to help us all do what we love in new ways. And the same innovation that goes into our products also applies to our practices — strengthening our commitment to leave the world better than we found it. This is where your work can make a difference in people’s lives. Including your own. Apple is an equal opportunity employer that is committed to inclusion and diversity. Visit apple.com/careers to learn more.