Nebius

Site Reliability Engineer in Hardware Infrastructure

Nebius
Apply
3 days ago
Amsterdam, NetherlandsMid Level

Responsibilities

  • Ensure fault tolerance, scalability, and uninterrupted operations for infrastructure services.
  • Use technology to solve infrastructure problems and optimize system performance.
  • Implement and improve CI/CD processes.
  • Troubleshoot complex hardware, software, and networking issues.
  • Build, roll out, and maintain internal hardware infrastructure platforms and tooling.

Requirements

  • Proficiency in Linux systems.
  • Expertise in Python and Bash scripting for automation.
  • Demonstrated ability to troubleshoot complex hardware, software, and networking problems.
  • Strong analytical and problem-solving skills focused on system performance optimization.
  • Working proficiency in English.
  • Experience designing, developing, and running high-load distributed systems is preferred.
  • Interest in backend development is preferred.

Benefits

  • Competitive compensation.
  • Career growth and learning opportunities.
  • Flexibility and ownership.
  • Collaborative and innovative culture.
  • Opportunity to work on impactful AI projects.
  • International environment with talented teams.
  • Applicants must be authorized to work in the country where they apply.

Tech Stack

Categories

DevOpsSite Reliability
Nebius

About Nebius

1,001-5,000 employees

The Nebius AI Cloud brings powerful full-stack infrastructure for AI developers and practitioners across startups, enterprises and science institutes to build and deploy generative AI applications and rapidly deliver scientific breakthroughs by training and running ML models within a secure, high-performance, and cost-optimized cloud environment.

Contact me