1 day ago
Toronto, CanadaSenior
Responsibilities
- Support and enhance mission-critical environments for new and existing cloud-native products and services.
- Improve monitoring, observability, reliability, performance, and service onboarding in collaboration with product development teams.
- Manage infrastructure deployment pipelines and troubleshoot onboarding and operational issues.
- Drive capacity planning, disaster recovery, configuration management, and platform readiness activities.
- Build automation and tools to reduce manual toil, improve developer experience, and increase system reliability.
- Define and manage SLOs and error budgets with engineering teams.
- Participate in incident, problem, and change management processes.
- Configure and integrate identity providers, authentication systems, and secure access protocols.
- Set up monitoring, logging, synthetic monitoring, distributed tracing, and vulnerability remediation.
- Provide rotational evening, weekend, and on-call operational support as needed.
Requirements
- Bachelor’s degree in Computer Science or a related field; a master’s degree is a plus.
- At least 3 years of experience in Site Reliability, DevOps, or Cloud Engineering roles.
- Expertise with Microsoft Azure and Infrastructure as Code using Bicep, ARM, and Terraform.
- Experience with Azure Monitor, Application Insights, DataDog, and Log Analytics.
- Hands-on experience onboarding and integrating identity providers such as Azure Entra ID, Okta, Keycloak, or PingFederate.
- Experience with SAML, OAuth, OIDC, OpenTelemetry, distributed tracing, Checkly, and Playwright or equivalent tools.
- Knowledge of AI/ML-based anomaly detection and log aggregation tools such as Microsoft Azure Anomaly Detector.
- Experience with Microsoft Defender Suite, Sentinel, Defender for Cloud, and KQL threat hunting.
- Understanding of networking, Kubernetes, Docker, APIs, PowerShell, Bash, Kusto, SQL, Cosmos DB, and PostgreSQL.
- Familiarity with SimCorp Dimension and Salesforce is a plus.
- Proficiency in IT service management practices focused on incident, change, and problem management.
- Experience managing both onboarding projects and live production operations.
- Ability to work collaboratively across engineering, product, client, vendor, and stakeholder teams.
Benefits
- Global hybrid work policy with two required office days per week and remote work on other days if desired.
- Inclusive and diverse company culture.
- Work-life balance focused on sustainable professional and personal responsibilities.
- Opportunities for professional development and individualized career growth.
- Health care, leave, retirement plans, and eligibility for an annual discretionary bonus.
Tech Stack
Categories
Site Reliability
About SimCorp
SimCorp builds SimCorp Dimension, a front-to-back investment management platform and managed services used by buy-side institutions such as asset managers, pension funds, and insurers. It sells software licenses and cloud-delivered operations (SaaS/managed services) covering portfolio management, trading, risk, accounting, and reporting. Founded in 1971 and headquartered in Copenhagen, it is a subsidiary of Deutsche Börse Group and reports serving 40 of the world’s top 100 financial companies.
