
Site Reliability Engineer - Hyderabad
InfraCloud Technologies3 months ago
Hyderābād, IndiaMid Level / Senior
Responsibilities
- Architect and automate Kubernetes solutions for air-gapped and multi-region clusters.
- Optimize CI/CD pipelines and cloud-native deployments.
- Troubleshoot Linux systems and improve observability using monitoring and logging tools.
- Support multi-region disaster recovery and high-availability infrastructure.
- Work with open-source projects and select appropriate infrastructure tools.
- Educate and guide teams on modern cloud-native infrastructure practices.
- Address scaling, security, and infrastructure automation challenges.
Requirements
- 4–6 years of relevant experience.
- Hands-on experience with Kubernetes, including air-gapped clusters.
- Experience managing on-premises servers and infrastructure.
- Strong Linux troubleshooting skills.
- Experience with Helm, Docker, Ingress, and Ingress Controllers.
- Strong expertise in automation, programmable infrastructure, and CI/CD pipelines.
- Proficiency with Prometheus, Grafana, ELK, and multi-region disaster recovery.
- Knowledge of security and compliance for regulated industries.
- Strong communication skills.
- Experience with GKE, RKE, Rook-Ceph, CKA, or CKAD is preferred.
Benefits
- Work mode is work-from-office five days per week.
- The role is based onsite in Hyderabad.
- Opportunity to work on high-impact Kubernetes projects in regulated industries.
- Opportunities for learning, open-source contributions, and innovation.
Tech Stack
Categories
DevOpsSite Reliability