Yotta Infrastructure

Senior DevOps Engineer

Yotta Infrastructure
Apply
5 months ago
Mumbai, IndiaSenior

Responsibilities

  • Design, build, and maintain scalable cloud and on-premise infrastructure across AWS, GCP, and Azure.
  • Implement Infrastructure as Code using Terraform, Ansible, or equivalent tools.
  • Manage Kubernetes clusters, container orchestration, node optimization, high availability, fault tolerance, and disaster recovery readiness.
  • Design and maintain CI/CD pipelines for backend, frontend, and AI workloads, including automated testing, security scanning, and deployment.
  • Implement blue-green, canary, and rolling deployment strategies while improving release velocity and stability.
  • Build observability systems for metrics, logs, and traces; define SLIs, SLOs, and SLAs for critical services.
  • Lead incident response, root-cause analysis, post-incident reviews, and reliability improvements.
  • Implement infrastructure, CI/CD, and runtime security practices, including secrets, access controls, and identity management.
  • Support SOC 2, ISO 27001, GDPR, and India DPDP Act compliance requirements.
  • Monitor and optimize compute, storage, networking, and AI infrastructure costs through auto-scaling, quotas, and cost-aware scheduling.
  • Create reusable infrastructure templates and DevOps practices, advise on scalability and deployment strategies, and mentor junior engineers.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
  • 5–8 years of experience in DevOps, Site Reliability Engineering, or Platform Engineering roles.
  • Proven experience operating production systems with uptime and SLA commitments.
  • Strong hands-on experience with Docker, Kubernetes, and container ecosystems.
  • Deep experience with AWS, GCP, or Azure.
  • Proficiency with Terraform, Ansible, Helm, or similar infrastructure tooling.
  • Experience with Git-based CI/CD systems such as GitHub Actions, GitLab CI, Jenkins, or Argo CD.
  • Experience with monitoring, logging, and tracing stacks.
  • Knowledge of networking, load balancing, and CDN architectures.
  • Strong scripting skills in Bash, Python, or an equivalent language.
  • Hands-on experience managing large-scale distributed systems.
  • Experience with SaaS, cloud platforms, or AI infrastructure is preferred.
  • Experience supporting AI/ML or GPU-heavy workloads, FinOps cost optimization, model serving patterns, or Next.js, Vercel, and Netlify-like pipelines is preferred.
  • Relevant AWS, GCP, Azure, or Kubernetes certifications are preferred but not mandatory.

Benefits

  • General day shift; the posting does not specify a rotational schedule.
  • Three interview rounds.
  • Opportunity to work on sovereign AI infrastructure, cloud platforms, and large-scale digital workloads in India.

Tech Stack

AnsibleArgo CDAWSAzureBashDatadogDockerGitHub ActionsGitLab CI/CDGoogle Cloud PlatformGrafanagRPCHelmJenkinsKubernetesNetlifyNext.jsPrometheusPythonTerraformVercel

Categories

Contact me