5 days ago
Bucharest, RomaniaSenior
Responsibilities
- Build and maintain a platform golden path for adding new microservices quickly.
- Improve, scale, and migrate CI/CD pipelines and shared pipeline abstractions.
- Enhance developer workstations and platform automation supporting an AI-capable client platform.
- Operate Kubernetes workloads, including scheduling, resource constraints, probes, disruption budgets, security contexts, ingress, and rollout troubleshooting.
- Create and maintain Helm charts or Kustomize bases and overlays used by multiple applications.
- Author and optimize Docker images using multi-stage builds, non-root images, build caching, and registries.
- Develop and maintain Terraform, Ansible, Python, and Bash automation.
- Improve dashboarding, alerting, observability, and service-level indicators, objectives, and alerting strategies.
- Participate in shared incident response, platform maintenance, and daily support for more than eight development teams.
- Automate recurring support requests and reduce operational toil.
- Support Kafka operations and Java, Spring Boot, and Angular delivery when needed.
Requirements
- At least 5 years of experience in DevOps, platform, or SRE roles, including real production ownership.
- In-depth Kubernetes expertise and the ability to debug failing rollouts.
- Experience designing Helm charts or building Kustomize bases and overlays used by multiple applications.
- Experience engineering CI/CD pipelines at scale; GitLab CI is strongly preferred.
- Experience creating shared, versioned pipeline abstractions used across many repositories.
- Experience authoring Docker images with multi-stage builds, small non-root images, build caching, and registries.
- Experience writing and maintaining Terraform and Ansible roles and modules.
- Strong automation skills in Python, Bash, or other programming languages, with clean, tested, reviewable code.
- Understanding of metrics, logs, traces, and application troubleshooting.
- Experience with Git, code review, trunk-adjacent branching, conventional commits, and small reviewable changes.
- Experience with HashiCorp Vault and Kubernetes secret-management patterns is strongly desirable.
- Experience authoring Prometheus and Grafana assets, including PromQL, recording rules, and dashboards as code, is strongly desirable.
- Experience with supply-chain security practices such as SBOM generation and CI checks is strongly desirable.
- Kafka operations experience, including topics, ACLs, and consumer lag, is strongly desirable.
- Enough SQL, Oracle, and MSSQL knowledge to assist during incidents is strongly desirable.
- Knowledge of Jenkins pipelines is strongly desirable.
- Experience supporting Java, Spring Boot, and Angular delivery is strongly desirable.
- Ability to work effectively with AI coding tools while reviewing their output for real infrastructure code.
- Experience defining and operating SLIs, SLOs, and alerting strategies is strongly desirable.
Benefits
- Hybrid role based in the Bucharest office.
- Strategic, innovation-focused project with strong business impact.
- Exposure to modern technologies and cloud-based architecture.
- Cross-border collaboration with teams in CER countries.
- Supportive and agile working environment with growth and development opportunities.
- Short Fridays.
- Private healthcare.
- Meal, vacation, and gift vouchers.
Tech Stack
AngularAnsibleApache KafkaBashDockerGitGitLab CI/CDGrafanaHelmJavaJenkinsKubernetesMicrosoft SQL ServerPrometheusPythonSpring BootSQLTerraform
