PayPal

Sr. Site Reliability Engineer

PayPal
Apply
16 hours ago
Chennai, IndiaSenior

Responsibilities

  • Act as incident commander with final decision authority during high-severity incidents, directing engineering teams and authorizing rollbacks, regional failovers, and emergency changes.
  • Review and troubleshoot Infrastructure as Code, Kubernetes manifests, CI/CD configurations, and cloud-native deployments during incidents and change approvals.
  • Drive site resilience, reliability, availability, scalability, capacity planning, monitoring, observability, automation, and disaster recovery initiatives.
  • Lead postmortems, root cause analysis, continuous improvement, and standardized incident-response playbooks and workflows.
  • Provide architectural direction, technical guidance, mentorship, and training to engineering teams and junior site reliability engineers.
  • Interface with executive leadership and cross-functional teams during critical incidents, postmortems, product rollouts, and change-management activities.

Requirements

  • At least 3 years of relevant experience with a bachelor’s degree, or an equivalent combination of education and experience.
  • At least 5 years of experience in site reliability engineering, infrastructure operations, or similar technical operations roles.
  • Significant hands-on experience with AWS or GCP; multi-cloud experience with AWS, GCP, or Azure is preferred.
  • Strong proficiency with Terraform, CloudFormation, Pulumi, Ansible, or equivalent infrastructure automation tools.
  • Significant hands-on experience with Kubernetes and CNCF ecosystem tools, including troubleshooting deployments, manifests, and cluster issues.
  • Ability to review Python, Go, and Bash code and YAML, HCL, and JSON configuration formats during incident troubleshooting.
  • Experience managing critical incidents in Infrastructure-as-Code-driven environments, including IaC state issues, GitOps failures, and cloud-native deployment problems.
  • Professional-level certification such as AWS Solutions Architect Professional or Google Cloud Professional Cloud Architect.
  • Experience with Splunk, Datadog, Prometheus, or Grafana and strong knowledge of networking, load balancing, CDN technologies, and DNS management.
  • Proficiency in Python, Bash, or PowerShell scripting, plus knowledge of distributed systems, microservices architecture, Docker, and Kubernetes.
  • Strong communication, analytical, problem-solving, collaboration, documentation, and high-pressure decision-making skills.

Benefits

  • PayPal provides comprehensive choice-based wellbeing programs, paid time off, healthcare coverage, financial-security resources, and mental-health support.
  • The role uses an alternating 12-hour shift pattern across three- to four-day periods.
  • PayPal’s balanced hybrid model generally includes three days in the office and two days either in the office or at home.

Tech Stack

AnsibleAWSAzureBashDatadogDockerGoGoogle Cloud PlatformGrafanaKubernetesPowerShellPrometheusPythonSplunkTerraform

Categories

Site Reliability
PayPal

About PayPal

10,000+ employees

PayPal builds digital payments platforms for consumers and merchants, including PayPal, Venmo, and Xoom, supporting online, in‑app, and in‑person checkout and peer‑to‑peer transfers. It earns primarily from transaction and merchant processing fees, plus value‑added services such as risk management and branded credit. Founded in 1998 and headquartered in San Jose, California, PayPal operates in roughly 200 markets and is a public company traded on NASDAQ.

Contact me