2 hours ago
Responsibilities
- Design, build, and operate reliable and scalable infrastructure in AWS and other cloud platforms.
- Evolve Kubernetes, Helm charts, Argo deployment systems, and platform guardrails for safe service deployment and operations.
- Improve observability using OpenMetrics and Datadog.
- Establish standards for Terraform infrastructure-as-code codebases.
- Create and support CI/CD pipelines, build tooling, and artifact repositories.
- Improve incident response tooling and lead incident response, post-mortems, and follow-up actions.
- Partner across organizational boundaries to deliver platform-wide infrastructure changes.
- Design and improve identity, networking, disaster recovery, business continuity, resilience, and cloud cost-efficiency practices.
Requirements
- 10+ years of software engineering, site reliability engineering, or infrastructure engineering experience with substantial production infrastructure experience.
- Deep cloud infrastructure knowledge and hands-on AWS experience.
- Strong infrastructure-as-code skills, preferably with Terraform, and experience creating safe, reusable infrastructure patterns.
- Solid software engineering skills and the ability to build automation and production tooling in a general-purpose programming language.
- Experience leading incident response and conducting blameless post-mortems or COE processes.
- Ability to make sound infrastructure decisions involving security, IAM, networking, and change management.
- Strong written and verbal communication skills for documenting designs, incidents, operational tradeoffs, and platform-wide changes.
- Ability to work independently in ambiguous environments and identify high-leverage problems.
- Preferred experience operating AWS at scale across multiple accounts, regions, and availability zones; GitOps and Argo CD; IAM and OIDC; disaster recovery and resilience testing; and cloud cost optimization.
Benefits
- Hybrid work arrangement with three days per week in the office.
- Latest hardware and software, including frontier AI models on day one.
- Daily catered lunches, breakfast, snacks, and soft drinks in the office.
- Home-office commuting costs covered for employees living outside Amsterdam.
- 25 vacation days based on full-time employment.
- Employer-paid collective health insurance with basic and additional packages.
- Defined pension contribution scheme.
- Equity program for team members.
- Employer-sponsored Employee Assistance Program through Aetna Resources for Living.
- Parental leave benefits for mothers and partners.
Tech Stack
Categories
DevOpsSite Reliability
About Flexport
Our mission is to make global trade so easy that there will be more of it.