2 days ago
Madrid, SpainSenior
Responsibilities
- Ensure the availability and reliability of managed hosting services for customer investment applications.
- Operate solutions securely, compliantly, and cost-effectively.
- Participate in the on-call pool and support high reliability, availability, and uptime.
- Troubleshoot infrastructure issues and improve resilience, availability, and scalability.
- Increase operational efficiency and help control cloud costs through automation.
- Provide hands-on guidance, mentorship, and collaborative problem-solving to peers.
Requirements
- At least 5 years of system engineering experience, including at least 3 years working with the listed platform technologies.
- Proficiency with Helm Charts, Kubernetes, Istio, ArgoCD, GitHub, and GitHub Actions.
- Understanding of cloud infrastructure security, cloud-native technologies, and desired state configuration.
- Knowledge of designing and implementing scalable, highly available cloud infrastructure.
- Automation-first mindset with a focus on making processes predictable and automated.
- Strong communication skills and good English proficiency for collaboration with international teams.
- Experience with monitoring and observability, particularly the Grafana stack and OpenTelemetry, is desirable.
- Knowledge of the Python stack is a plus.
- Mentorship skills and the ability to provide hands-on technical guidance.
