1 day ago
Responsibilities
- Mentor and support other site reliability engineers.
- Define requirements during the product lifecycle and influence new designs and standards.
- Deploy and maintain internal platforms and tools.
- Partner with multiple teams to improve product and service availability, reliability, scalability, and usability.
- Improve the Compute Cloud Interface platform for faster error detection and remediation, performance, and reliability.
- Develop automation that supports daily operations and reduces toil.
- Participate in on-call rotations and guide restoration and repair of service-impacting issues.
- Troubleshoot and resolve internal escalations and customer-impacting incidents.
Requirements
- Five years of relevant experience.
- Bachelor’s degree in Computer Science or equivalent experience.
- Experience automating with Python, Golang, and Bash.
- Understanding of systems reliability, observability, monitoring, and SLO adherence.
- Experience with SaltStack, Terraform, Ansible, and Jenkins.
- Hands-on mastery of Linux administration and container-based platforms such as Docker.
- Experience with Prometheus, Grafana, Loki, nginx, Envoy, HAProxy, and Redis.
Benefits
- Akamai provides benefits supporting employees’ health, well-being, finances, and life beyond work.
- Flexible work arrangements are available at home, in an office, or through a combination of both.
Categories
Site Reliability
About Akamai
Akamai creates and sells healthy, 100% natural, nutrient rich, radically simplified personal care for the mindful consumer via our website (only).