
Senior Cloud Platform Engineer
Deluxe Corporation9 days ago
Remote, GermanySenior
Responsibilities
- Design, build, automate, and maintain cloud infrastructure for AI services, media workflows, internal platforms, and enterprise applications.
- Develop and maintain infrastructure-as-code using Terraform, AWS CDK, CloudFormation, or similar tools.
- Support AWS networking, compute, storage, IAM, security controls, monitoring, environment automation, and cost management.
- Build scalable infrastructure for production services, batch processing, high-throughput workloads, and GPU-enabled environments.
- Create and maintain CI/CD pipelines, deployment automation, configuration management, and operational tooling.
- Support Linux systems, Dockerized applications, runtime environments, service deployments, and production release processes.
- Implement standards for reliability, observability, security, backup and recovery, access management, incident response, and operational readiness.
- Troubleshoot complex issues across infrastructure, systems, applications, networking, security, and deployment pipelines.
- Develop reusable provisioning, configuration, monitoring, deployment, support, and environment-management patterns.
- Maintain operational documentation, architecture diagrams, runbooks, and knowledge-transfer materials.
- Reduce manual work through scripting, automation, platform improvements, and standardized tooling.
Requirements
- Strong professional experience in cloud infrastructure, DevOps, systems engineering, site reliability engineering, or platform engineering.
- Hands-on experience designing and operating AWS infrastructure in production environments.
- Experience with infrastructure-as-code tools such as Terraform, AWS CDK, CloudFormation, or similar technologies.
- Strong knowledge of AWS networking, VPCs, IAM, security groups, load balancing, compute, storage, and monitoring services.
- Strong Linux administration, scripting, automation, and troubleshooting skills.
- Experience with Docker, containerized workloads, CI/CD pipelines, deployment automation, and production release practices.
- Experience supporting production systems with monitoring, logging, alerting, incident response, and operational runbooks.
- Understanding of cloud security, identity and access management, secrets management, and least-privilege access patterns.
- Ability to troubleshoot issues across cloud, systems, network, application, and deployment layers.
- Strong written and verbal communication skills.
- Preferred experience with GPU workloads, AI/ML environments, model-serving platforms, high-performance computing, or data-intensive systems.
- Preferred experience with Kubernetes, ECS, EKS, Slurm, HPC clusters, or distributed compute environments.
- Preferred experience with Ansible, Helm, GitHub Actions, GitLab CI, Jenkins, or similar tooling.
- Preferred experience with CloudWatch, Prometheus, Grafana, Datadog, ELK/OpenSearch, or OpenTelemetry.
- Experience with media, localization, content-processing, workflow automation, or high-throughput production platforms is preferred.
- Familiarity with security hardening, vulnerability management, compliance-driven operations, and enterprise infrastructure standards is preferred.
- AWS certifications are a plus.
Benefits
- Hybrid work arrangement in Aachen, Germany.
- The role supports AI platforms, media workflows, internal platforms, and enterprise application environments within a global media and entertainment services company.
Tech Stack
AnsibleAWSDatadogDockerGitHub ActionsGitLab CI/CDGrafanaHelmJenkinsKubernetesLinuxPrometheusTerraform