GrepJob
GoDaddy

Staff Site Reliability Engineer-Observability

GoDaddy
Apply
about 3 hours ago
Delhi, IndiaStaff+
H1B Sponsor

Responsibilities

  • Build and evolve observability using Prometheus/Mimir and Grafana.
  • Steer cloud migration efforts to maintain service reliability.
  • Own patching compliance and vulnerability remediation at scale.
  • Operate and extend a multi-tenant Kubernetes/ArgoCD platform.
  • Contribute to AI/automation initiatives to reduce operational toil.
  • Consult with partner dev teams on metrics and monitoring standards.
  • Mentor other engineers and help mature operational practices.

Requirements

  • 8+ years of hands-on experience with AWS.
  • 5+ years of experience with Kubernetes.
  • 4+ years of expertise in Linux administration.
  • Strong coding skills in languages such as Go, Python, or Ruby.
  • Experience in coding infrastructure as code using Terraform or Ansible.

Benefits

  • Paid time off and retirement savings options.
  • Bonus/incentive eligibility and equity grants.
  • Participation in employee stock purchase plan.
  • Competitive health benefits and family-friendly perks.

Tech Stack

AnsibleAWSGoGradleGrafanaJenkinsKubernetesLinuxMavenMySQLPostgreSQLPrometheusPythonRubySQLTerraform