
Senior DevOps Engineer
Yotta Infrastructure5 months ago
Mumbai, IndiaSenior
Responsibilities
- Design, build, and maintain scalable cloud and on-premise infrastructure across AWS, GCP, and Azure.
- Implement Infrastructure as Code using Terraform, Ansible, or equivalent tools.
- Manage Kubernetes clusters, container orchestration, node optimization, high availability, fault tolerance, and disaster recovery readiness.
- Design and maintain CI/CD pipelines for backend, frontend, and AI workloads, including automated testing, security scanning, and deployment.
- Implement blue-green, canary, and rolling deployment strategies while improving release velocity and stability.
- Build observability systems for metrics, logs, and traces; define SLIs, SLOs, and SLAs for critical services.
- Lead incident response, root-cause analysis, post-incident reviews, and reliability improvements.
- Implement infrastructure, CI/CD, and runtime security practices, including secrets, access controls, and identity management.
- Support SOC 2, ISO 27001, GDPR, and India DPDP Act compliance requirements.
- Monitor and optimize compute, storage, networking, and AI infrastructure costs through auto-scaling, quotas, and cost-aware scheduling.
- Create reusable infrastructure templates and DevOps practices, advise on scalability and deployment strategies, and mentor junior engineers.
Requirements
- Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
- 5–8 years of experience in DevOps, Site Reliability Engineering, or Platform Engineering roles.
- Proven experience operating production systems with uptime and SLA commitments.
- Strong hands-on experience with Docker, Kubernetes, and container ecosystems.
- Deep experience with AWS, GCP, or Azure.
- Proficiency with Terraform, Ansible, Helm, or similar infrastructure tooling.
- Experience with Git-based CI/CD systems such as GitHub Actions, GitLab CI, Jenkins, or Argo CD.
- Experience with monitoring, logging, and tracing stacks.
- Knowledge of networking, load balancing, and CDN architectures.
- Strong scripting skills in Bash, Python, or an equivalent language.
- Hands-on experience managing large-scale distributed systems.
- Experience with SaaS, cloud platforms, or AI infrastructure is preferred.
- Experience supporting AI/ML or GPU-heavy workloads, FinOps cost optimization, model serving patterns, or Next.js, Vercel, and Netlify-like pipelines is preferred.
- Relevant AWS, GCP, Azure, or Kubernetes certifications are preferred but not mandatory.
Benefits
- General day shift; the posting does not specify a rotational schedule.
- Three interview rounds.
- Opportunity to work on sovereign AI infrastructure, cloud platforms, and large-scale digital workloads in India.
Tech Stack
AnsibleArgo CDAWSAzureBashDatadogDockerGitHub ActionsGitLab CI/CDGoogle Cloud PlatformGrafanagRPCHelmJenkinsKubernetesNetlifyNext.jsPrometheusPythonTerraformVercel