Mirantis

Senior Site Reliability Engineer (SRE)

Mirantis
Apply
4 hours ago
Remote, Poland or Warsaw, PolandSenior

Responsibilities

  • Develop, implement, maintain, and troubleshoot cloud and AI infrastructure based on open-source software.
  • Deploy AI infrastructure on NVIDIA-certified hardware according to engineering architecture and implementation designs.
  • Optimize infrastructure performance, reliability, scalability, security, networking, and storage.
  • Troubleshoot, debug, and resolve complex technical issues across hardware and software environments.
  • Design and implement AI-driven automation across the DevOps lifecycle, including code development and maintenance.
  • Collaborate with distributed international teams, stakeholders, and customers to refine requirements and solve technical challenges.
  • Participate in code reviews and process improvements to maintain quality standards.
  • Lead technical tasks, make independent technical decisions, and mentor team members and customers.
  • Facilitate knowledge transfer during customer delivery phases.
  • Stay current with cloud operations, development, and open-source industry practices.
  • Travel internationally up to 25% when needed.

Requirements

  • Bachelor's degree in Computer Science or a related field, or equivalent experience.
  • At least 5 years of DevOps, software development, or similar professional experience.
  • Strong experience with cloud and infrastructure technologies, including Kubernetes and/or OpenStack.
  • Experience with high-performance data center processing, networking, and storage.
  • Exposure to Golang and working knowledge of Python and JavaScript.
  • Knowledge of distributed systems, microservices architecture, and CI/CD pipelines.
  • Strong troubleshooting and debugging skills across networking, storage, Linux, and Kubernetes, with attention to performance and security.
  • Ability to lead technical tasks and collaborate with geographically distributed teams.
  • Ability to work directly with customers and make independent judgment calls with limited day-to-day oversight.
  • Excellent written, spoken, and customer-facing communication skills.
  • Preferred experience includes network or storage architecture, high-performance computing, GPU infrastructure, GPU scheduling, MIG/vGPU, RDMA/RoCE or InfiniBand fabrics, NVLink, DCGM health-checking, GPU driver or firmware lifecycle, or NVIDIA AI Enterprise.
  • Preferred qualifications include open-source community participation, upstream contributions, conference presentations, and experience with Rancher, OpenShift, or VMware.

Benefits

  • Professional development and training.
  • Opportunities to attend conferences and working groups.
  • Company outings, happy hours, hackathons, and tech talks.
  • Competitive compensation package with a strong benefits plan.
  • Work with open-source cloud infrastructure technologies and Fortune 500 and Global 2000 customers.
  • International travel may be required up to 25%.
  • Salary range of 70000-85000 Euro per year.

Tech Stack

AWSGoJavaScriptKubernetesLinuxOpenShiftOpenStackPythonRancher

Categories

DevOpsSite Reliability
Mirantis

About Mirantis

501-1,000 employees

Mirantis builds Kubernetes-native infrastructure software and services for enterprises to run cloud, container, and AI workloads across on-premises, public cloud, and edge environments. It sells subscriptions, support, and managed services around products like Mirantis Kubernetes Engine and the Lens Kubernetes platform; the company acquired Docker Enterprise in 2019. Founded in 1999 and headquartered in Campbell, California, Mirantis is privately held and an IREN company serving global enterprises.

Contact me