GrepJob
NICE

Senior Cloud Site Reliability Engineer

NICE
Apply
about 4 hours ago
London, United KingdomSenior / Mid Level
H1B Sponsor

Responsibilities

  • Monitor availability and system health of the production environment.
  • Build software and systems to manage platform infrastructure and applications.
  • Improve reliability, quality, and time-to-market of software solutions.
  • Measure and optimize system performance to meet customer needs.
  • Provide operational support for large distributed software applications.
  • Gather and analyze metrics for performance tuning and fault finding.
  • Partner with development teams to enhance services through testing.
  • Participate in system design consulting and capacity planning.
  • Create sustainable systems through automation.

Requirements

  • 3-6 years of experience in systems engineering, automation, and reliability.
  • Proficiency in at least one programming language and experience with scripting languages.
  • Deep understanding of cloud computing platforms and their reliability constraints.
  • Experience with infrastructure as code tools like CloudFormation or Terraform.
  • Strong knowledge of CI/CD concepts and tools.
  • Experience with containerization technologies and microservices architecture.
  • Familiarity with monitoring and observability tools.
  • Excellent problem-solving skills for troubleshooting complex issues.
  • Experience in incident management and driving incident response efforts.

Benefits

  • Hybrid work model with 2 days in the office and 3 days remote.
  • Collaborative office environment focused on teamwork and innovation.

Tech Stack

Amazon DynamoDBAnsibleAWSBashC#ChefCircleCIDatadogDockerGitLab CI/CDGoGrafanaJavaJenkinsKubernetesPowerShellPrometheusPuppetPythonSplunkTerraform

Categories