
Senior Platform Reliability Engineer - Azure
London Stock Exchange Group3 days ago
Hyderābād, IndiaSenior
Responsibilities
- Build, deploy, administer, and optimize Azure cloud infrastructure across compute, storage, networking, security, and Kubernetes services.
- Manage Azure Virtual Machines, Storage Accounts, Virtual Networks, NSGs, Load Balancers, VPN Gateways, ExpressRoute, and AKS.
- Implement infrastructure as code using Terraform and ARM Templates and develop automation to reduce operational overhead.
- Build and support CI/CD pipelines using Azure DevOps and GitLab Actions.
- Improve service reliability, performance, availability, capacity planning, and operational efficiency across the service lifecycle.
- Implement monitoring, alerting, logging, and observability using Azure Monitor, Log Analytics, and related tools.
- Maintain platform security using RBAC, Azure Policies, Key Vault, governance controls, and compliance processes.
- Build and support high-availability, backup, disaster recovery, and business continuity solutions.
- Conduct incident management, root cause analysis, problem management, and production performance tuning.
- Develop operational scripts, tools, software components, runbooks, documentation, and automated responses using PowerShell, Python, Bash, or other languages.
- Collaborate with application development, architecture, security, and operations teams on resilient cloud solutions.
- Participate in system design reviews, resource planning, launch readiness assessments, and operational excellence initiatives.
- Mentor junior engineers and define engineering standards and continuous improvement initiatives.
Requirements
- Bachelor's degree or equivalent experience in computer science, software engineering, electronics/electrical engineering, or an equivalent technical field.
- 7-10 years of overall IT infrastructure, cloud engineering, DevOps, or site reliability engineering experience.
- At least 4 years of hands-on Microsoft Azure experience.
- Strong hands-on experience with Azure compute, storage, networking, security and governance, AKS, Terraform, ARM Templates, and CI/CD implementation.
- Experience with PowerShell, Python, and/or Bash scripting and Windows and Linux system administration.
- Experience supporting large-scale production environments and critical applications.
- Expertise in automation, infrastructure provisioning, monitoring, reliability engineering, observability, high availability, disaster recovery, and performance optimization.
- Experience implementing enterprise-scale cloud governance and security controls.
- Strong understanding of cloud architecture, distributed systems, and reliability engineering principles.
- Preferred Azure certifications include AZ-104, AZ-305, AZ-400, or equivalent.
- Experience with containerization, Kubernetes operations, and cloud-native architectures.
- Strong analytical, problem-solving, interpersonal, and cross-functional collaboration skills.
Benefits
- Healthcare benefits, retirement planning, paid volunteering days, and wellbeing initiatives.
- Collaborative and creative culture with opportunities for innovation, knowledge sharing, and mentoring.
- Global work environment at an equal-opportunity employer with sustainability and charitable engagement programs.
About London Stock Exchange Group
London Stock Exchange Group (LSEG) provides market infrastructure, financial data, and analytics to banks, asset managers, and corporations. Its businesses span capital markets (London Stock Exchange), post-trade clearing (LCH), and information services including FTSE Russell indices and the Refinitiv data platform, funded by transaction fees, subscriptions, and licensing. Headquartered in London and publicly listed on the LSE, the group serves global customers seeking trading, clearing, benchmarks, and enterprise data solutions.