Marks & Spencer Group plc

Systems Reliability Engineer

Marks & Spencer Group plc
Apply
2 hours ago
Tokyo, JapanSenior

Responsibilities

  • Design, build, and maintain systems in close collaboration with engineering and development teams.
  • Troubleshoot issues across hardware, software, applications, and networks.
  • Identify and implement automation for platform deployment, management, and service visibility.
  • Proactively identify and address systems reliability risks.
  • Collaborate with global and regional team members on a follow-the-sun basis.
  • Represent the Reliability & Production Engineering organization in design reviews and operational readiness exercises.

Requirements

  • At least 6 years of experience.
  • Demonstrated ability to troubleshoot problems, debug systems, and identify root causes.
  • Hands-on experience with AppDynamics, Grafana, Splunk, and Dynatrace.
  • Experience with Ansible, GitHub, or other automation, configuration-management, or release-management tools.
  • Automation experience using scripting languages such as Python, Bash, Perl, or Ruby.
  • One higher-level programming language is desired.
  • Awareness of modern software and systems architectures, including load balancing, databases, queueing, caching, distributed-systems failure modes, microservices, and cloud environments.
  • Practical experience running large-scale systems is advantageous.

Benefits

  • Global collaboration with regional teams in a follow-the-sun operating model.
  • Morgan Stanley offers comprehensive employee benefits and perks and supports employees and their families throughout their work-life journey.
  • Equal opportunity employment with a stated commitment to diversity and inclusion.

Tech Stack

AnsibleBashGrafanaPerlPythonRubySplunk

Categories

Site Reliability
Marks & Spencer Group plc

About Marks & Spencer Group plc

10,000+ employees
Contact me