
Systems Reliability Engineer
Marks & Spencer Group plc2 hours ago
Tokyo, JapanSenior
Responsibilities
- Design, build, and maintain systems in close collaboration with engineering and development teams.
- Troubleshoot issues across hardware, software, applications, and networks.
- Identify and implement automation for platform deployment, management, and service visibility.
- Proactively identify and address systems reliability risks.
- Collaborate with global and regional team members on a follow-the-sun basis.
- Represent the Reliability & Production Engineering organization in design reviews and operational readiness exercises.
Requirements
- At least 6 years of experience.
- Demonstrated ability to troubleshoot problems, debug systems, and identify root causes.
- Hands-on experience with AppDynamics, Grafana, Splunk, and Dynatrace.
- Experience with Ansible, GitHub, or other automation, configuration-management, or release-management tools.
- Automation experience using scripting languages such as Python, Bash, Perl, or Ruby.
- One higher-level programming language is desired.
- Awareness of modern software and systems architectures, including load balancing, databases, queueing, caching, distributed-systems failure modes, microservices, and cloud environments.
- Practical experience running large-scale systems is advantageous.
Benefits
- Global collaboration with regional teams in a follow-the-sun operating model.
- Morgan Stanley offers comprehensive employee benefits and perks and supports employees and their families throughout their work-life journey.
- Equal opportunity employment with a stated commitment to diversity and inclusion.
Tech Stack
Categories
Site Reliability