
Senior Site Reliability Engineer
Lloyds Banking Group1 hour ago
Hyderābād, IndiaSenior / Staff+
Responsibilities
- Own one or more areas of cloud infrastructure resources and supervise SRE work in those areas.
- Improve observability and prioritize operational service improvements to improve SLOs.
- Manage significant aspects of the data management system and support its operation and development.
- Apply incident management, distributed-systems, resilience, disaster-recovery, automation, and infrastructure-as-code practices.
- Develop the capabilities of direct reports and provide specialized training or coaching across the organization.
- Identify and deliver improvements to IT security, operational, change-management, risk-management, and strategic-planning processes.
- Analyze technical problems, define solutions, develop product specifications, and design testing procedures and standards.
Requirements
- 8–11 years of experience.
- Experience with Google Cloud, Harness, and Kubernetes.
- Strong SRE fundamentals, including SLOs and SLIs, observability, and incident management.
- Expertise in distributed systems, resilience and disaster-recovery design, automation-first practices, infrastructure-as-code or programming, and regulatory awareness in high-availability environments.
- Ability to manage significant technical areas, supervise or develop others, and deliver outcomes within established systems and processes.
Benefits
- Hybrid working and flexible working options are supported.
- The posting end date is Tuesday 29 September 2026.
Tech Stack
Categories
Site Reliability