2 months ago
Remote, United KingdomStaff+

Responsibilities

  • Lead and grow a small DevOps engineering team supporting Unite Services and Unite Integrations.
  • Own platform reliability, security, operational efficiency, deployments, releases, and multi-environment support.
  • Operate and evolve cloud and on-prem Kubernetes clusters, including upgrades, networking, ingress, scaling, and security.
  • Design and maintain GitHub-based CI/CD pipelines, self-hosted runners, repository migrations, branching strategies, and GitOps practices.
  • Manage supporting infrastructure including Redis, Consul, Harbor, Chart Museum, API gateways, and message queues.
  • Drive security hardening, vulnerability remediation, WAF and rate-limiting controls, secrets management, and compliance initiatives.
  • Define and execute disaster recovery and ransomware recovery strategies and maintain recovery runbooks.
  • Own observability, dashboards, alerts, incident response, post-incident improvements, and platform documentation.
  • Contribute to Internal Developer Platform capabilities, developer experience improvements, RFCs, and platform maturity models.

Requirements

  • Strong background in platform or DevOps engineering for complex distributed systems.
  • Deep understanding of compute, networking, storage, DNS, certificates, and load balancing.
  • Deep production Kubernetes and Docker experience, including cluster operations, upgrades, scaling, security, ingress, and RBAC.
  • Experience with at least one major cloud provider; Azure is preferred and AWS or GCP is a plus.
  • Experience working in hybrid cloud and on-prem environments.
  • Strong experience designing and operating CI/CD pipelines, including GitHub Actions and self-hosted runners.
  • Experience with infrastructure-as-code tools such as Terraform, Crossplane, and Helm.
  • Strong scripting or coding skills in Python, Go, Bash, or PowerShell.
  • Experience with configuration-management tools such as Ansible or Chef.
  • Experience with Prometheus, Grafana, and centralized logging systems such as ELK or Splunk.
  • Understanding of infrastructure, network, and application security practices; WAF, rate limiting, Qualys, and CrowdStrike experience is a plus.
  • Proven experience leading DevOps or platform teams, mentoring engineers, and owning outcomes from design through ongoing operations.
  • Experience leading small or medium-sized projects using Agile, Scrum, or similar methods.
  • Experience with Internal Developer Platforms, API gateways, Apache Flink, ransomware preparedness, disaster recovery, business continuity, or multi-region and multi-tenant systems is preferred.

Benefits

  • All candidates must be located in the United Kingdom.
  • Intermedia promotes from within and emphasizes teamwork, transparency, accountability, and an inclusive work environment.

Tech Stack

AnsibleApache FlinkBashChefCloudflareConsulDockerGitHub ActionsGoGrafanaHelmKubernetesPowerShellPrometheusPythonRabbitMQRedisSplunkTensorFlowTerraform

Categories

Intermedia Intelligent Communications

About Intermedia Intelligent Communications

1,001-5,000 employees
Contact me