over 1 year ago
Delhi, IndiaStaff+
Responsibilities
- Design and develop infrastructure-as-code components for AWS-hosted products and internal engineering tools.
- Build and deploy a scalable, secure security SaaS platform using automated and repeatable CI/CD processes.
- Administer Linux systems at scale through automation.
- Secure infrastructure using TLS, bastion hosts, certificate management, authentication and authorization, and network segmentation.
- Design, develop, and manage deployments across multiple Kubernetes clusters.
- Manage, maintain, and monitor infrastructure across cloud and on-premises environments.
- Implement consistent logging, monitoring, and diagnostic systems.
- Participate in on-call duties.
Requirements
- 8+ years of experience in DevOps or platform engineering.
- 4+ years of experience setting up AWS infrastructure for SaaS product development organizations.
- Strong knowledge of AWS infrastructure and services including load balancers, IAM, KMS, EC2, CloudWatch, CloudTrail, and Lambda.
- 4+ years of experience building infrastructure with Terraform.
- 3+ years of experience with Kubernetes and Helm.
- Strong knowledge of configuration-management systems such as Ansible.
- Experience with CI/CD code-management and deployment technologies such as GitLab and Docker.
- Familiarity with Nginx, HAProxy, Kafka, and other public-cloud components.
- Experience with Grafana, Prometheus, or the LGTM stack is a plus.
- MLOps experience is a plus.
- Experience deploying and managing Aurora RDS-Postgres, ElastiCache, Cassandra, OpenSearch or Elasticsearch, and ClickHouse is a plus.
- Experience working with current AI technologies for infrastructure management is a plus.
- Strong time-management, prioritization, communication, and adaptability skills.
Tech Stack
AnsibleApache CassandraApache KafkaAWSClickHouseDockerElasticsearchGrafanaHelmKubernetesLinuxPostgreSQLPrometheusTerraform
