11 days ago
Bengaluru, IndiaSenior
Responsibilities
- Design, build, and evolve CI/CD pipelines for frequent, reliable, and automated software releases.
- Build and maintain infrastructure as code for AWS environments using Terraform or similar technologies.
- Deploy and operate production workloads using Kubernetes and cloud-native technologies.
- Improve monitoring, logging, tracing, alerting, troubleshooting, incident response, root-cause analysis, and operational reliability.
- Partner with engineering teams to improve reliability, scalability, performance, deployment quality, and production readiness.
- Strengthen security and resilience through IAM, secrets management, cloud governance, high availability, backup, and disaster recovery practices.
- Build reusable tooling and self-service capabilities that improve developer productivity.
- Provide hands-on technical leadership, mentoring, documentation, and practical guidance across engineering teams.
Requirements
- 8+ years of experience in DevOps, cloud infrastructure, platform engineering, SRE, or related engineering roles.
- Strong hands-on experience building and operating production workloads on AWS.
- Deep experience with CI/CD and release automation using GitFlow, GitHub Actions, ArgoCD, Jenkins, or similar technologies.
- Strong infrastructure-as-code experience, particularly with Terraform, CloudFormation, or equivalent tools.
- Hands-on experience with Kubernetes, Docker, Helm, and containerized production environments.
- Strong scripting and automation skills using Python, Bash, PowerShell, or similar languages.
- Experience implementing observability across metrics, logs, and distributed tracing using Prometheus, Grafana, OpenTelemetry, Datadog, or equivalent platforms.
- Understanding of networking, DNS, certificates, IAM, secrets management, Linux, and cloud security fundamentals.
- Ability to diagnose complex production and infrastructure problems and drive them through resolution.
- Experience operating multi-tenant SaaS platforms at scale is preferred.
- Experience building internal developer platforms, self-service infrastructure, reusable deployment patterns, or developer tooling is preferred.
- Exposure to cloud security automation, vulnerability management, compliance controls, and governance is preferred.
- Experience designing for high availability, disaster recovery, resilience, and production operability is preferred.
- AWS certification or equivalent depth of practical AWS experience is preferred.
