Apple

Site Reliability Engineer (SRE), Observability, London

Apple
Apply
15 hours ago
London, United KingdomSenior

Responsibilities

  • Design, operate, and scale observability infrastructure across geographically distributed data centers.
  • Manage and improve telemetry, logging, monitoring, provisioning, configuration management, and software deployment systems.
  • Design, author, and release production code in Go, Python, or comparable languages.
  • Automate manual operations and iteratively improve engineering processes.
  • Deploy, support, and monitor services, platforms, and application stacks.
  • Perform scale testing, disaster recovery, capacity planning, and distributed-system troubleshooting.

Requirements

  • Experience managing and scaling distributed systems in public, private, or hybrid cloud environments.
  • Ability to design, author, and release code in languages such as Go or Python.
  • Understanding of the Linux operating system and standard networking protocols and components.
  • Experience managing large numbers of diverse systems with configuration-management or software-delivery platforms such as Puppet and Spinnaker is preferred.
  • Experience with service deployment, support, and monitoring, plus scale testing, disaster recovery, and capacity planning, is preferred.
  • Familiarity with microservices architecture and Kubernetes container orchestration is preferred.
  • Strong ownership, integrity, communication, and collaboration skills.

Tech Stack

Categories

Site Reliability
Apple

About Apple

10,000+ employees

Apple designs and sells consumer electronics, software, and services for consumers and professionals worldwide, including iPhone, Mac, iPad, Apple Watch, and AirPods, plus platforms like iOS/macOS and services such as the App Store, iCloud, Music, and TV+. Its business combines device sales with services and subscriptions and in-house silicon design. Founded in 1976, Apple is headquartered in Cupertino, California, and trades on NASDAQ as AAPL.

Contact me