Nebius

Senior Site Reliability Engineer

Nebius
Apply
9 months ago
Remote, Worldwide or Amsterdam, NetherlandsSenior

Responsibilities

  • Ensure fault tolerance, scalability, and uninterrupted operations for the service.
  • Use cloud technology to solve infrastructure problems.
  • Implement and improve CI/CD processes.

Requirements

  • Solid experience with programming languages such as Go, Python, or C++.
  • Strong understanding of classic algorithms and data structures.
  • Commercial experience with and deep understanding of Unix systems and network technology.
  • Experience with containerization and configuration-management systems including Ansible, Salt, Terraform, Docker, Kubernetes, and Helm.
  • Experience with backend development is a bonus.
  • Experience designing, developing, and running high-load distributed systems is a bonus.
  • Commercial experience with a variety of cloud platforms is a bonus.
  • Participation in coding interviews is required.

Benefits

  • Competitive salary and comprehensive benefits package.
  • Opportunities for professional growth within Nebius.
  • Flexible working arrangements.
  • Dynamic and collaborative work environment.
  • Global work environment with R&D hubs across Europe, North America, and Israel.

Categories

DevOpsSite Reliability
Nebius

About Nebius

1,001-5,000 employees

The Nebius AI Cloud brings powerful full-stack infrastructure for AI developers and practitioners across startups, enterprises and science institutes to build and deploy generative AI applications and rapidly deliver scientific breakthroughs by training and running ML models within a secure, high-performance, and cost-optimized cloud environment.

Contact me