2 hours ago
London, United KingdomSenior
Responsibilities
- Architect, scale, and maintain observability backends.
- Extend and maintain OpenTelemetry collectors and SDKs.
- Build scalable telemetry ingestion and routing pipelines.
- Contribute to Golden Path SDKs and auto-instrumentation.
- Ensure Kubernetes workloads are observable and resilient.
- Embed observability standards across platform and application teams.
- Improve incident response through better telemetry coverage.
- Provide industry observability expertise and input to the long-term roadmap.
Requirements
- Deep experience with observability stacks and operating or onboarding SaaS observability platforms at scale, such as Datadog, New Relic, or Dynatrace.
- Strong hands-on experience with OpenTelemetry.
- Familiarity with public cloud infrastructure, ideally AWS.
- Proficiency in Kubernetes and DevOps tooling such as Terraform, Argo CD, Helm, and Jenkins.
- Experience with metrics, logs, and tracing backends.
- Scripting proficiency in Go, Python, or a similar language.
- An industry background in observability or site reliability engineering.
- Desirable experience with profiling technologies including eBPF, Pixie, or Parca; synthetic monitoring; AI-observability tools; and Kafka.
Benefits
- Highly competitive compensation plus an annual discretionary bonus.
- Lunch provided through Just Eat for Business and access to a dedicated barista bar.
- 35 days of annual leave.
- 9% company pension contributions.
- Informal dress code and excellent work/life balance.
- Comprehensive healthcare and life assurance.
- Cycle-to-work scheme.
- Monthly company events.
- Inclusive recruitment experience with accommodations available for applicants with disabilities or special needs.
