9 hours ago
Bengaluru, IndiaSenior
Responsibilities
- Engage with product teams to design and implement resilient, scalable infrastructure solutions.
- Operate, monitor, and triage production and non-production environments.
- Collaborate on code, infrastructure, design reviews, and process improvements.
- Evaluate and integrate technologies that improve reliability, security, and performance.
- Develop automation to provision, configure, deploy, and monitor Apple services.
- Participate in on-call rotations and provide technical expertise during service-impacting events.
- Contribute to capacity planning, scale testing, and disaster recovery exercises.
Requirements
- 6+ years of demonstrated expertise in Site Reliability Engineering, infrastructure operations, or a DevOps-focused role.
- Understanding of SRE principles including monitoring, alerting, error budgets, fault analysis, capacity planning, automation, and toil reduction.
- Proficiency in at least one of Python, Go, or Java.
- Experience managing and scaling distributed systems in public, private, or hybrid cloud environments.
- Experience with microservices architecture and Kubernetes or similar container orchestration technologies.
- BS or MS in Computer Science or a related field, or equivalent work experience.
- Experience running Tier 1 services with 24/7 support.
- Strong understanding of Linux fundamentals, networking principles, and system management.
- Strong ownership, communication, and collaboration skills.
Tech Stack
Categories
Site Reliability
About Apple
Apple designs and sells consumer electronics, software, and services for consumers and professionals worldwide, including iPhone, Mac, iPad, Apple Watch, and AirPods, plus platforms like iOS/macOS and services such as the App Store, iCloud, Music, and TV+. Its business combines device sales with services and subscriptions and in-house silicon design. Founded in 1976, Apple is headquartered in Cupertino, California, and trades on NASDAQ as AAPL.
