
Staff Site Reliability Engineer
ServiceNow17 hours ago
Remote, Ireland or Dublin, IrelandStaff+
Responsibilities
- Design, build, and operate cloud-native engineering platforms for software validation, release validation, and production readiness.
- Maintain production-like release and test ServiceNow environments that improve deployment readiness and release confidence.
- Build automated test pipelines, observability, reliability signals, deployment intelligence, and quality gates into CI/CD workflows.
- Develop automation, reusable frameworks, self-service environments, test data management, mock services, and developer productivity tooling.
- Design and enhance Kubernetes platforms supporting scalable test infrastructure, release automation, cloud-native workloads, and developer self-service.
- Implement automated validation for failure detection, deployment verification, policy enforcement, security checks, resilience testing, and operational health.
- Resolve complex platform, infrastructure, networking, software engineering, and automation challenges.
- Partner with engineering teams to improve platform reliability, release quality, cloud-native adoption, and engineering practices.
- Participate in architecture reviews, technical design discussions, and implementation of scalable automation-first solutions.
- Mentor engineers through technical guidance, code reviews, knowledge sharing, and best practices.
Requirements
- 8+ years of experience in SRE, DevOps, Platform Engineering, Software Engineering, or Infrastructure Engineering with a bachelor's degree; alternatively, 6 years with a master's degree, 3 years with a PhD, or equivalent experience.
- Hands-on Kubernetes experience covering cluster operations, networking, storage, security, autoscaling, and multi-cluster environments.
- Experience building and operating scalable, highly available cloud-native platforms.
- Experience integrating Kubernetes with CI/CD, GitOps, automated test pipelines, deployment validation, and cloud-native deployment workflows.
- Experience with developer productivity automation, progressive delivery, canary deployments, feature flags, automated rollback, and deployment verification.
- Experience with chaos engineering, resilience testing, disaster recovery, reliability validation, observability, monitoring, SLI/SLOs, incident management, and distributed-system operations.
- Strong software engineering experience designing, developing, testing, and debugging applications using Python, Go, Java, or Ruby.
- Experience with test automation frameworks, test orchestration, regression testing, test impact analysis, flaky test detection, parallel execution, test data management, service virtualization, contract testing, and synthetic testing.
- Experience with DevOps automation, CI/CD, GitOps, Infrastructure as Code, configuration management, and Kubernetes ecosystem tools.
- Experience operating Kubernetes across AWS, Azure, and Google Cloud is preferred.
- Experience with AI-assisted engineering, intelligent testing, operational automation, or cloud-native engineering platforms is preferred.
- Ability to solve complex technical problems, independently drive projects, collaborate across globally distributed engineering teams, and mentor others.
Benefits
- Regular employee position with a flexible or remote work persona in the EMEA region.
- ServiceNow provides an accessible and inclusive workplace and reasonable accommodations for candidates who need them.
Tech Stack
Categories
DevOpsSite Reliability
About ServiceNow
ServiceNow builds a cloud platform for enterprise digital workflows, covering IT service management, customer service, HR service delivery, security operations, and operations management, plus tools for custom app development. It sells subscription SaaS to large organizations and public-sector agencies to automate processes and connect data across systems. Founded in 2004 and headquartered in Santa Clara, California, ServiceNow is a public company listed on the NYSE under the ticker NOW.