2 months ago
Responsibilities
- Design and build end-to-end automation pipelines for cluster provisioning and new region and environment builds, including AWS account bootstrapping for sovereign regions.
- Translate manual web-based infrastructure workflows into reliable, idempotent, API-driven automation.
- Provision and validate monitoring, alerting, telemetry pipelines, and network circuits for new regions.
- Integrate automation with GitLab CI, Backstage, and release controls with pipeline gates and exit criteria.
- Build agents and orchestration layers for incident intelligence, failure-pattern detection, remediation recommendations, change-risk scoring, and production-readiness gates.
- Develop automated rollbacks based on business signals such as error rate, latency, and revenue impact.
- Build and maintain Model Context Protocol server integrations with ServiceNow, Jira, Confluence, and AWS APIs using appropriate guardrails.
- Collaborate with SRE, DevOps, Systems, Network, and Voice Engineering teams to identify and automate high-toil workflows.
- Design solutions that preserve independent operation and compliant data boundaries for UK, AU, and EU sovereign instances.
- Contribute to the DevOps AI Working Group and deliver measurable reductions in manual effort within the first 90 days.
Requirements
- At least five years of software engineering experience focused on backend systems, infrastructure automation, or platform engineering.
- Production-grade experience with Python or Go; TypeScript or Node.js experience is a plus.
- Deep AWS experience, including IAM, EC2, EKS, RDS, Lambda, CloudWatch, and multi-account architectures.
- Hands-on experience building automation against REST or GraphQL APIs.
- Familiarity with GitLab CI or GitHub Actions and infrastructure-as-code tools such as Terraform, CDK, or Pulumi.
- Experience building AI-powered workflows with LLM APIs such as Anthropic or OpenAI, including prompt engineering, tool use, function calling, and agentic orchestration.
- Strong understanding of distributed systems, Kubernetes, and cloud-native observability involving metrics, logs, and traces.
- Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent work experience.
- Experience with agentic AI frameworks such as LangGraph, CrewAI, AutoGen or AG2, Semantic Kernel, or Smolagents is a bonus.
- Experience in regulated or sovereign cloud environments, Backstage and CI/CD integration, observability platforms such as OpenObserve or Grafana, incident automation, Kafka, or internal infrastructure tooling is a bonus.
Tech Stack
Categories
About NICE
NiCE is transforming the world with AI that puts people first. Our purpose-built AI-powered platforms automate engagements into proactive, safe, intelligent actions, empowering individuals and organizations to innovate and act, from interaction to resolution. Trusted by organizations throughout 150+ countries worldwide, NiCE’s platforms are widely adopted across industries connecting people, systems, and workflows to work smarter at scale, elevating performance across the organization, delivering proven measurable outcomes.