
Principal Platform Engineer
ISS Governance11 days ago
London, United KingdomStaff+
Responsibilities
- Shape the architecture, evolution, and operational excellence of global cloud platforms across GCP and AWS.
- Design and mature AWS landing zones aligned with GCP standards for governance, automation, security, and developer experience.
- Develop platform services, self-service capabilities, CI/CD tooling, observability solutions, and automation frameworks.
- Build platform capabilities, tooling, operational patterns, and guardrails for AI-powered applications, agentic workflows, and emerging AI technologies.
- Provide hands-on technical leadership during critical incidents and resolve high-impact platform engineering problems.
- Evaluate emerging technologies, develop proofs of concept, and convert successful innovations into reusable production capabilities.
- Partner with engineering, architecture, security, and product teams to define technical direction and engineering standards.
- Drive improvements in reliability, observability, automation, security, cost efficiency, and developer experience.
- Mentor engineers and promote technical excellence, continuous learning, and platform engineering best practices.
- Contribute to the long-term evolution of the Platform and AI Engineering function.
- Participate in out-of-hours support and major incident escalation when required.
Requirements
- At least 10 years of experience in platform engineering, cloud engineering, or related disciplines, including large-scale enterprise platform design and operation.
- Strong experience designing and implementing cloud-native platforms on GCP and/or AWS.
- Experience building or maturing cloud landing zones, governance frameworks, and self-service developer platforms.
- Deep understanding of containers, Kubernetes, serverless technologies, APIs, networking, security, and identity management.
- Strong Infrastructure as Code expertise with master proficiency in Terraform.
- Experience developing and operating large-scale CI/CD, automation, and platform engineering capabilities.
- Experience evaluating emerging technologies, developing proofs of concept, and delivering production-ready platform capabilities.
- Practical experience enabling AI and data-driven workloads on cloud platforms, including related tooling, operations, and governance.
- Strong software engineering and automation skills, including expertise in Python and reusable platform tooling.
- Experience with GitHub Actions, Apigee, Airflow, and related cloud-native technologies.
- Expertise with observability platforms such as Datadog, Prometheus, Grafana, ELK, Splunk, or equivalent, including monitoring, logging, tracing, reliability engineering, and incident management.
- Sound technical judgment and the ability to balance innovation, risk, operational excellence, and business outcomes.
- Ability to influence technical direction, build consensus, and drive adoption across multiple teams.
- Experience coaching and mentoring engineers in global, distributed, or multinational organizations.
- Excellent communication, documentation, and stakeholder engagement skills.
- Bachelor's or Master's degree in Computer Science, Engineering, or a related discipline, or equivalent practical experience.
- Nice-to-have experience with AI-powered applications, machine learning workloads, agentic systems, GCP AI services, vector databases, retrieval-augmented generation, model operations, AI governance, or multi-cloud standards.
Benefits
- The role is part of a global organization operating across 20 countries with a culture focused on diversity, inclusion, creativity, innovation, and career growth.
- The position includes out-of-hours support and major incident escalation responsibilities when required.
Tech Stack
Apache AirflowAWSDatadogGitHub ActionsGoogle Cloud PlatformGrafanaKubernetesPrometheusPythonSplunkTerraform