8 months ago
Barcelona, SpainSenior
Responsibilities
- Own the availability and performance of mission-critical services and build automation to prevent recurring problems.
- Improve system scalability, observability, and alerting.
- Build platform tooling that accelerates software development and enables engineers to ship safely and quickly.
- Practice sustainable incident response and conduct blameless postmortems.
- Collaborate with product teams on technical issues and new system designs.
- Improve the quality and credibility of the team’s technical execution.
- Apply FinOps practices to infrastructure decisions.
- Participate in the team’s paid on-call rotation.
Requirements
- Approximately eight years of experience designing, building, maintaining, and troubleshooting high-traffic distributed systems.
- Hands-on expertise using agentic AI to enhance engineering workflows, establish quality gates and safety nets, and propose AI-native alternatives while applying judgment to architecture, business logic, and security.
- Proficient software engineering skills, with Python preferred.
- Experience with at least one major public cloud and its services and infrastructure, preferably AWS.
- Experience implementing observability and alerting.
- Willingness to participate in paid on-call rotations.
- Business orientation and a data-driven approach.
- Strong communication skills and minimum C1 English proficiency.
Benefits
- Open, collaborative, dynamic, and diverse company culture.
- Monthly allowance for Preply lessons, a Learning & Development budget, and time off for self-development.
- Financial package including equity, leave allowance, and health insurance.
- Relocation package for candidates joining the Barcelona Hub from elsewhere.
- Access to free mental health support platforms.
- Access to Gympass-partnered wellness and gym centers throughout Spain.
