1 month ago
London, United KingdomSenior
Responsibilities
- Own the availability and performance of mission-critical services and build automation to prevent recurring problems.
- Improve platform scalability, observability, alerting, and infrastructure efficiency.
- Build tooling and core platform services that improve developer productivity and accelerate software delivery.
- Design, build, maintain, and troubleshoot high-traffic distributed systems and storage systems.
- Practice sustainable incident response and conduct blameless postmortems.
- Collaborate with product teams on technical issues and new system designs.
- Apply FinOps practices and participate in the team’s paid on-call rotation.
- Use agentic AI thoughtfully in engineering workflows, including quality gates, safety nets, and AI-native alternatives.
Requirements
- Approximately 8 years of experience designing, building, maintaining, and troubleshooting high-traffic distributed systems.
- Proficient software engineering skills, with Python preferred.
- Experience with at least one major public cloud and its services and infrastructure, preferably AWS.
- Experience implementing observability and alerting.
- Willingness to participate in paid on-call rotations.
- Ability to use agentic AI beyond code generation while applying judgment to architecture, business logic, and security.
- Business-oriented, data-driven approach and strong communication skills.
- Minimum C1 English level.
Benefits
- Generous monthly allowance for lessons on Preply.com.
- Learning and Development budget, including time off for self-development.
- Competitive financial package with equity and leave allowance.
- Opportunity to impact learners and tutors across more than 175 countries.
- The role includes a paid on-call rotation.
