2 months ago
Ontario, CA, USAMid Level / Senior

Responsibilities

  • Design, deploy, and manage scalable infrastructure across AWS, GCP, and Azure.
  • Build repeatable infrastructure deployments using Pulumi and Python, with CloudFormation and Terraform as additional tools.
  • Manage and scale containerized applications using AWS ECS and EKS (Kubernetes).
  • Build and maintain CI/CD pipelines with GitHub Actions.
  • Administer and scale RabbitMQ, Apache Kafka, and AWS SQS messaging systems.
  • Configure and optimize Elasticsearch and OpenSearch clusters.
  • Migrate legacy Java workloads running on JBOSS, WebLogic, or similar application servers into containerized cloud environments.
  • Implement monitoring, alerting, and logging frameworks using New Relic, Grafana, Prometheus, and CloudWatch.
  • Implement DevSecOps practices including container scanning, dependency alerts, and automated secrets management.
  • Drive multi-cloud cloud cost optimization, reduce alert noise, improve runbooks, and decrease mean time to resolution.

Requirements

  • 3–5 years of professional experience as a DevOps, Site Reliability (SRE), or Cloud Engineer deploying applications into modern cloud environments.
  • Deep AWS expertise and production-level experience with at least one additional cloud provider, GCP or Azure.
  • Strong proficiency with Pulumi and Python for infrastructure automation.
  • Production experience managing workloads with AWS ECS and EKS.
  • Advanced knowledge of GitHub Actions for workflow automation and deployment pipelines.
  • Hands-on experience scaling and troubleshooting AWS SQS, Kafka, and RabbitMQ.
  • Experience configuring and tuning Elasticsearch and/or OpenSearch.
  • Experience migrating legacy application-server workloads such as JBOSS or WebLogic into containerized environments.
  • Strong experience creating dashboards and alerts in New Relic.
  • Experience with Grafana and Prometheus is preferred.
  • Experience implementing DevSecOps practices and cloud cost optimization across multi-cloud environments.
  • Strong problem-solving, automation, communication, and collaboration skills.

Tech Stack

Apache KafkaAWSAzureElasticsearchGitHub ActionsGoogle Cloud PlatformGrafanaJavaKubernetesPrometheusPythonRabbitMQTerraform

Categories

Warner Music Group

About Warner Music Group

5,001-10,000 employees
Contact me