2 months ago
Responsibilities
- Operate and support enterprise compute platforms across hardware, operating systems, virtualization, and container orchestration layers.
- Develop automation for provisioning, patching, upgrades, validation, and routine platform operations.
- Deploy and support Ubuntu systems, hypervisors, Kubernetes clusters, and related platform services.
- Implement and maintain PXE provisioning environments using Redfish APIs for large-scale server deployments.
- Operate and support KVM on Ubuntu, OpenStack environments, and Harvester HCI.
- Monitor performance, capacity, and availability while proactively addressing reliability risks.
- Troubleshoot complex issues across hardware, operating systems, virtualization, OpenStack, and Kubernetes.
- Manage SLAs, KPIs, and error budgets and participate in on-call escalation support.
- Collaborate globally on change management, documentation, and operational best practices.
- Create and maintain runbooks, operational procedures, and technical documentation.
- Mentor junior engineers and support a culture of technical learning and excellence.
Requirements
- At least six years of experience as a DevOps Engineer, Site Reliability Engineer, or Infrastructure Operations Engineer focused on compute.
- Strong hands-on experience operating bare metal compute environments at scale.
- Experience with PXE boot, automated OS provisioning, server imaging, and Redfish-based Bare Metal as a Service platforms.
- Strong Linux administration skills, especially with Ubuntu.
- Experience building and supporting infrastructure and platform automation pipelines.
- Proficiency with Infrastructure as Code tools such as Terraform and Ansible.
- Operational experience with KVM on Ubuntu, OpenStack, and Harvester HCI.
- Experience deploying and operating production Kubernetes environments.
- Strong scripting skills in Python, Bash, or similar languages.
- Understanding of toil reduction, error budgets, SLAs, troubleshooting, and root cause analysis in distributed systems.
- Bachelor’s degree in computer science or equivalent professional experience.
- Preferred qualifications include CIS/NIST security knowledge, ITIL certifications, Ubuntu certifications, CKA or CKS certification, and experience in telco, edge cloud, or large enterprise environments.
- A master’s degree in computer science, IT, engineering, or a related field and relevant industry certifications are preferred.
Benefits
- Collaborative team focused on infrastructure excellence.
- Complex technical challenges involving scalable infrastructure solutions.
- Opportunity to shape a next-generation private cloud platform.
- Access to current tools, frameworks, and upstream project developments.
- Role is based in India; on-call escalation support is required.
About Five9
Five9 provides a comprehensive suite of CX solutions, powered by Five9 Genius AI, to elevate customer experiences that deliver better business outcomes in the era of The New CX. The New CX redefines how brands connect with customers through seamless and efficient AI-driven journeys that anticipate and meet each customer’s unique needs. Our unified cloud-native offering enables AI and human agents to create hyper-personalized customer experiences, so every customer interaction is more connected, effortless, and personal. Trusted by 3,000+ customers and 1,400+ partners globally, Five9 brings together the power of our AI, our platform, and our people to drive AI-elevated CX.