3 hours ago
Bengaluru, IndiaMid Level
Responsibilities
- Automate deployment and operations for Kong’s Managed Gateways across various cloud environments.
- Monitor system health, performance, and uptime while working toward 99.99% availability.
- Resolve complex production incidents and participate in on-call rotations.
- Build resilient tools and systems that improve platform reliability and operational efficiency.
- Collaborate with engineering teams to design, review, and implement resilient, scalable services.
- Help prevent technical debt and support sustainable operations as the company grows.
Requirements
- At least 2 years of experience applying Site Reliability Engineering principles in a production environment.
- Proficiency in Golang or Python for automation, tooling, and infrastructure as code.
- Hands-on experience with Kubernetes and a major cloud platform such as AWS, GCP, or Azure.
- Familiarity with monitoring, logging, and alerting tools such as Prometheus, Grafana, or Datadog.
- Understanding of networking concepts, distributed systems, and API gateways.
- Experience with Kong Gateway or other API management platforms is preferred.
- Relevant cloud certifications, such as AWS Certified DevOps Engineer or Kubernetes Administrator, are preferred.
- Active contributions to open-source projects or developer communities are preferred.
Tech Stack
Categories
DevOpsSite Reliability
