about 3 hours ago
Base Salary
$133k - $209k/yr
Responsibilities
- Drive initiatives for best practices in data streaming, processing, and monitoring.
- Deploy and manage services on Kubernetes platforms like Amazon EKS and GKE.
- Provision and manage cloud infrastructure using Terraform.
- Maintain and optimize CI/CD pipelines with Jenkins, ArgoCD, and GitHub Actions.
- Work with cloud-native data services such as AWS Kinesis and Google Dataflow.
- Develop automation scripts using Python to support DevOps processes.
- Monitor system performance and troubleshoot issues.
- Implement SRE practices to enhance service reliability and cost-effectiveness.
- Collaborate with cross-functional teams to improve development workflows.
Requirements
- 7+ years of experience in DevOps, Site Reliability Engineering, or Cloud Infrastructure.
- Strong experience with AWS and GCP data services.
- Proficiency in deploying workloads on Kubernetes in production.
- Hands-on experience with Infrastructure-as-Code using Terraform.
- Expertise in CI/CD pipeline management.
- Programming skills in Python for automation.
- Experience with observability and monitoring tools.
- Strong understanding of SRE principles.
- Experience with cloud cost optimization strategies.
- Ability to work collaboratively in an agile environment.
Benefits
- Comprehensive benefits package.
- Holistic mind, body, and lifestyle programs for overall well-being.
Tech Stack
Apache AirflowApache FlinkApache KafkaApache SparkDatadogGitHub ActionsGoogle BigQueryGrafanaIstioJenkinsKubernetesPrometheusPythonRabbitMQSnowflakeTerraform