Zscaler

Staff Site Reliability Engineer (Linux/Network troubleshooting/Scripting)

Zscaler
Apply
2 hours ago
Bengaluru, IndiaStaff+
H1B Sponsor

Responsibilities

  • Design, implement, and manage advanced cloud management automation to reduce toil and accelerate delivery.
  • Oversee and optimize containerized architectures using EKS and GKE for production performance.
  • Create, deploy, and optimize scalable monitoring and alerting systems.
  • Own cloud operations, deployments, on-call support, and incident management.
  • Design and tune Linux- and BSD-based systems for availability, security, and resilience.
  • Contribute to cross-functional technology projects and advise on feasibility of new initiatives.
  • Improve reliability by preventing recurring incidents through process, monitoring, and architectural enhancements.
  • Contribute to OS and software packaging and distribution and mentor teammates on SRE practices.

Requirements

  • At least 4 years of relevant experience designing, analyzing, and troubleshooting large-scale distributed systems.
  • Demonstrated curiosity about and active exploration of AI tools to improve workflows and problem-solving.
  • Deep hands-on experience with SRE practices, Python, Terraform, Ansible, networking, Kubernetes, and AWS.
  • Strong understanding of web security and HTTP, SSL/TLS, DNS, SQL, and networking fundamentals.
  • Experience building observability architectures and complex dashboards and managing Grafana.
  • Understanding of SLIs, SLOs, and error budgets.
  • Strong DevOps experience across delivery pipelines, source control management, builds, releases, and continuous integration tools or frameworks.
  • Preferred: advanced knowledge of virtualization, cloud architecture, cloud services, and automated deployment methods.
  • Preferred: experience resolving critical escalations and preventing recurring incidents.
  • Preferred: experience with OS and software packaging and distribution and mentoring others on SRE best practices.

Benefits

  • Hybrid working model.
  • Various health plans.
  • Vacation and sick time-off plans.
  • Parental leave options.
  • Retirement options.
  • Education reimbursement.
  • In-office perks.
  • Inclusive workplace and reasonable support or accommodations during recruiting.

Tech Stack

Categories

Site Reliability
Zscaler

About Zscaler

5,001-10,000 employees

Zscaler (NASDAQ: ZS) is a pioneer and global leader in zero trust security. The world’s largest businesses, critical infrastructure organizations, and government agencies rely on Zscaler to secure users, branches, applications, data & devices, and to accelerate digital transformation initiatives. Distributed across 160+ data centers globally, the Zscaler Zero Trust Exchange™ platform combined with advanced AI combats billions of cyber threats and policy violations every day and unlocks productivity gains for modern enterprises by reducing costs and complexity. Stay Connected: LinkedIn: https://www.linkedin.com/company/zscaler Twitter: https://www.twitter.com/zscaler Facebook: https://www.facebook.com/Zscaler/ Instagram: https://www.instagram.com/zscalerinc/

Contact me