1 day ago
Toronto, CanadaSenior
Responsibilities
- Build and operate continuous and progressive delivery capabilities for services running on Azure Kubernetes Service.
- Own shared CI/CD workflows, reusable Terraform modules, GitOps reconciliation, canary and blue/green deployments, and service mesh delivery.
- Create self-service platform capabilities, automated guardrails, reference implementations, standards, and documentation for engineering teams.
- Design and test highly available, multi-region Azure solutions, including automated failover and disaster-recovery capabilities.
- Integrate observability, SLI/SLOs, release gating, rollback capabilities, and service-behavior evaluation into delivery processes.
- Establish secure and compliant delivery paths using Azure Policy, Kubernetes admission controls, and workload identity.
- Design and deliver AI-assisted platform tooling for onboarding, configuration validation, access auditing, drift detection, and incident triage.
- Incorporate approval gates, checkpointing, and auditability into AI-assisted platform solutions.
Requirements
- At least 8 years of experience in software, platform, or infrastructure engineering, including ownership of an end-to-end platform capability used by other teams.
- At least 8 years of enterprise Azure experience, including Entra ID, workload identity, networking, governance, Azure Policy, Key Vault, and managed data and messaging services.
- At least 5 years of experience operating production Kubernetes across multiple clusters and regions, including a service mesh such as Istio or an equivalent.
- At least 5 years supporting progressive delivery through canary or blue/green deployment approaches.
- At least 4 years of Terraform and CI/CD experience, including reusable modules adopted by other teams.
- At least 4 years developing reusable GitHub Actions workflows with approval gating and OpenID Connect federation with Azure.
- At least 2 years of production experience with GitOps technologies such as Flux CD or Argo CD.
- Production experience with Python, Bash, YAML, and containers.
- At least 2 years of hands-on observability and policy-as-code experience, including New Relic or Datadog, SLI/SLO development, and Kubernetes admission-control policies.
- Current hands-on experience building large-language-model solutions using agent frameworks, tool calling, or the Model Context Protocol for infrastructure or operational challenges.
- Experience delivering technology solutions in regulated enterprise environments with formal change-management, control, and audit requirements.
- Preferred experience includes chaos engineering and disaster-recovery testing, Model Context Protocol servers or agentic workflows with state management and human approval gates, service mesh operations at scale, and API management configuration promotion.
Benefits
- Flexible work environment with support for learning, career growth, well-being, and inclusion.
- Hybrid working arrangement in Toronto, Ontario.
- Eligible employees may receive health, dental, mental health, vision, disability, life and AD&D, adoption/surrogacy, wellness, and employee/family assistance benefits.
- Eligible employees may participate in retirement savings plans, pension, global share ownership with employer matching, and financial education and counseling resources.
- Paid holidays, vacation, personal and sick days, and statutory leaves of absence are available in Canada.
- Employees may participate in incentive programs tied to business and individual performance.
Tech Stack
Categories
About Manulife
Manulife is a Toronto‑headquartered public financial services company that provides life and health insurance, retirement plans, and wealth and asset management to individuals and institutions. It earns premiums and fee income from insurance, investment, and advisory products delivered across Canada, Asia, and Europe, and operates as John Hancock in the United States. Founded in 1887, the company is listed on the Toronto Stock Exchange.
