1 hour ago
Paris, FranceMid Level
Responsibilities
- Build and maintain infrastructure-as-code, CI/CD pipelines, and zero-downtime deployment capabilities.
- Design and implement observability covering metrics, logs, traces, alerting, and service-level objectives.
- Establish and operate the on-call rotation and author and maintain actionable runbooks.
- Harden infrastructure security in collaboration with the security lead.
- Maintain platform stability and availability for a live 24/7 regulated service.
- Use AI-assisted tools for infrastructure-as-code and pipeline authoring with human review before deployment.
- Apply AI to anomaly detection, alert triage, incident summarization, and post-mortem drafting.
- Deploy, monitor, and operate in-product AI/ML services, including rollback procedures.
Requirements
- Strong experience with infrastructure-as-code, including Terraform, containers, Kubernetes, and CI/CD pipelines.
- Proven experience with 24/7 observability, incident response, and SLO/error-budget practices.
- Experience building secure, auditable infrastructure for regulated workloads.
- Ability to remain calm under pressure, take ownership, and produce clear, actionable runbooks.
- Familiarity with AI-assisted engineering practices, AIOps, incident summarization, and monitoring AI/ML services.
- Strong analytical, problem-solving, and communication skills.
- Degree in computer science, engineering, or a related field.
- At least three years of relevant experience.
Benefits
- Based in Paris, Porto, Milan, or Rome.
- International, diverse, and innovative team environment.
- Continuous learning and real responsibility for professional development.
- Opportunity to contribute to digital asset infrastructure and responsible capital markets.
- Start date is as soon as possible.
Tech Stack
Categories
DevOpsSite Reliability