2 hours ago
Hyderābād, IndiaStaff+
Responsibilities
- Act as the technical lead for SRE initiatives across multiple Product Areas.
- Drive strategic use of Microsoft Azure across onboarding and site reliability disciplines.
- Architect scalable, secure, and automated solutions for client onboarding and live operations.
- Design and evolve cross-cutting platform capabilities including observability, CI/CD pipelines, Infrastructure as Code standards, and disaster recovery frameworks.
- Govern Azure implementation patterns for standardization, reliability, and cost efficiency.
- Solve complex reliability challenges involving distributed cloud systems.
- Advise engineering leads and product owners on cloud platform decisions, trade-offs, and risk mitigation.
- Collaborate with Information Security, Platform Engineering, and Architecture teams on compliance and cloud controls.
- Define SLOs, SLIs, and other reliability metrics across departments.
- Lead root cause analysis, major incident postmortems, and reliability retrospectives.
- Mentor and coach senior and lead engineers and build SRE communities of practice.
- Represent the SRE function in executive planning, roadmap definition, and technical due diligence.
- Contribute to SimCorp’s transformation into a SaaS-first, cloud-native company.
Requirements
- Bachelor’s or master’s degree in Computer Science, Engineering, or a related field.
- 10+ years of experience in Site Reliability Engineering, Cloud Infrastructure, or Platform Architecture roles.
- Extensive expertise in Microsoft Azure, including architecture, deployment, automation, and cost optimization.
- Extensive knowledge of Windows Servers OS and Windows-based desktop application troubleshooting.
- Strong knowledge of enterprise system operations and system administration across complicated system landscapes.
- Strong understanding of cloud-native and hybrid architectures, distributed systems, networking, and security.
- Mastery of Infrastructure as Code using Terraform, ARM, Bicep, and related tooling.
- Deep knowledge of observability stacks including Azure Monitor, Log Analytics, Grafana, and Application Insights.
- Experience leading complex incident and problem management efforts at scale.
- Broad technical skills including Kubernetes, Docker, CI/CD pipelines, SQL, APIs, and scripting.
- Strong foundation in ITIL processes and operational excellence.
- Ability to influence senior stakeholders, lead through ambiguity, and align engineering with business needs.
- Experience in regulated, security-conscious environments such as financial services.
- Demonstrated commitment to mentorship, knowledge sharing, and engineering culture.
- Ability to think strategically while delivering pragmatic, hands-on solutions.
- Experience with Citrix, AD, WAC/WSUS is a plus.
Benefits
- Attractive salary, executive-level bonus scheme, and pension are offered as part of the work agreement.
- Flexible working hours are available.
- Hybrid work models are available.
- Professional growth is individualized to the employee’s leadership journey.
- Access is provided to strategic initiatives and innovation programs across the organization.
