SimCorp

Principal Site Reliability Engineer

SimCorp
Apply
2 hours ago
Hyderābād, IndiaStaff+

Responsibilities

  • Act as the technical lead for SRE initiatives across multiple Product Areas.
  • Drive strategic use of Microsoft Azure across onboarding and site reliability disciplines.
  • Architect scalable, secure, and automated solutions for client onboarding and live operations.
  • Design and evolve cross-cutting platform capabilities including observability, CI/CD pipelines, Infrastructure as Code standards, and disaster recovery frameworks.
  • Govern Azure implementation patterns for standardization, reliability, and cost efficiency.
  • Solve complex reliability challenges involving distributed cloud systems.
  • Advise engineering leads and product owners on cloud platform decisions, trade-offs, and risk mitigation.
  • Collaborate with Information Security, Platform Engineering, and Architecture teams on compliance and cloud controls.
  • Define SLOs, SLIs, and other reliability metrics across departments.
  • Lead root cause analysis, major incident postmortems, and reliability retrospectives.
  • Mentor and coach senior and lead engineers and build SRE communities of practice.
  • Represent the SRE function in executive planning, roadmap definition, and technical due diligence.
  • Contribute to SimCorp’s transformation into a SaaS-first, cloud-native company.

Requirements

  • Bachelor’s or master’s degree in Computer Science, Engineering, or a related field.
  • 10+ years of experience in Site Reliability Engineering, Cloud Infrastructure, or Platform Architecture roles.
  • Extensive expertise in Microsoft Azure, including architecture, deployment, automation, and cost optimization.
  • Extensive knowledge of Windows Servers OS and Windows-based desktop application troubleshooting.
  • Strong knowledge of enterprise system operations and system administration across complicated system landscapes.
  • Strong understanding of cloud-native and hybrid architectures, distributed systems, networking, and security.
  • Mastery of Infrastructure as Code using Terraform, ARM, Bicep, and related tooling.
  • Deep knowledge of observability stacks including Azure Monitor, Log Analytics, Grafana, and Application Insights.
  • Experience leading complex incident and problem management efforts at scale.
  • Broad technical skills including Kubernetes, Docker, CI/CD pipelines, SQL, APIs, and scripting.
  • Strong foundation in ITIL processes and operational excellence.
  • Ability to influence senior stakeholders, lead through ambiguity, and align engineering with business needs.
  • Experience in regulated, security-conscious environments such as financial services.
  • Demonstrated commitment to mentorship, knowledge sharing, and engineering culture.
  • Ability to think strategically while delivering pragmatic, hands-on solutions.
  • Experience with Citrix, AD, WAC/WSUS is a plus.

Benefits

  • Attractive salary, executive-level bonus scheme, and pension are offered as part of the work agreement.
  • Flexible working hours are available.
  • Hybrid work models are available.
  • Professional growth is individualized to the employee’s leadership journey.
  • Access is provided to strategic initiatives and innovation programs across the organization.

Tech Stack

Categories

DevOpsSite Reliability
SimCorp

About SimCorp

1,001-5,000 employees
Contact me