
Staff Site Reliability Engineer - Paze
Early Warning Services2 hours ago
Tokyo, JapanStaff+
Base Salary
$131k - $192k/yr
Responsibilities
- Design and implement software, tools, and patterns that improve application and infrastructure availability, scalability, performance, and latency.
- Build automation and tooling for deployments, configuration changes, application management, and disaster recovery.
- Design and promote observability and monitoring systems to detect problems and identify root causes.
- Evaluate application capacity, identify performance bottlenecks, and recommend scaling approaches.
- Lead or support high-priority incident resolution, troubleshoot complex technical issues, and improve the root-cause analysis process.
- Serve as a technical liaison for applications and provide documentation and runbooks to Level 1 and Level 2 teams.
- Partner with application development teams on software lifecycle requirements, microservice design patterns, and reliability practices.
- Develop reusable standards and patterns, provide senior technical expertise, and coach and mentor team members.
- Participate in a 24/7 on-call rotation.
Requirements
- Bachelor’s degree in Business, Computer Science, or a related field, with typically 8+ years of related progressive experience in a technical or software development environment, including post-graduate degree experience.
- Proven ability to lead teams through high-priority incidents and improve root-cause analysis processes.
- Excellent troubleshooting skills and experience resolving technical issues in complex environments.
- Hands-on experience designing and developing with one or more of Python, Go, or Java.
- Experience with Docker, microservices architecture, messaging frameworks, database technologies, caching layers, Linux administration, and CI/CD pipeline implementation.
- Experience with Git, Chef, Maven, Jenkins, TCP/UDP/IP protocols, and leading cross-functional technical teams.
- Proven track record designing and building complex end-to-end systems, including full-stack development.
- Preferred experience with Java, Ruby, Python, JavaScript, Go, AWS, Docker, Kubernetes, and Swarm.
- Experience supporting applications in a 24/7 customer-facing production environment.
- Successful completion of a background and drug screen and ability to perform essential functions with or without reasonable accommodation.
Benefits
- Hybrid work model for positions located in Scottsdale, San Francisco, Chicago, or New York.
- Medical, dental, and vision coverage, plus company contributions to an HSA and pre-tax FSA options.
- 401(k) plan with a 100% company Safe Harbor Match on the first 6% deferred immediately upon eligibility.
- Flexible time off for exempt employees, generous PTO for non-exempt employees, 11 paid company holidays, and one paid volunteer day.
- 12 weeks of paid parental leave.
- Maven Family Planning support for fertility, adoption, surrogacy, pregnancy, postpartum, pediatrics, and returning to work.
Tech Stack
Amazon DynamoDBApache KafkaAWSChefDockerGitGoJavaJavaScriptJenkinsKubernetesLinuxMavenPythonRedisRuby
Categories
Site Reliability