7 months ago
Responsibilities
- Lead technical initiatives that improve the reliability, performance, and scalability of critical systems.
- Design and implement resilient, scalable solutions using distributed-systems expertise.
- Lead complex troubleshooting, identify root causes, and implement advanced optimizations.
- Develop and champion automation frameworks and tools, potentially guiding junior engineers.
- Design and implement monitoring and logging solutions that provide actionable insights across teams.
- Lead incident response and provide technical direction and mentorship during incidents.
- Lead and contribute to in-depth blameless post-mortems and drive improvements from their findings.
- Mentor and guide SREs and engineers in their technical development.
Requirements
- At least five years of experience in SRE, platform engineering, or software development with a strong operational focus.
- Experience providing technical leadership, guidance, or mentorship to engineering teams.
- Expert practical knowledge of cloud platforms, especially GCP.
- Hands-on experience with Kubernetes and infrastructure-as-code tools including Terraform, Helm, and ArgoCD.
- Strong command of Python, Go, and Bash.
- Expertise building and using monitoring and observability tools including Prometheus, Grafana, and the ELK stack.
- Senior-level analytical, problem-solving, and debugging skills.
- Strong communication, collaboration, and influencing skills.
- Organizational and leadership skills are a plus.
- Educational credentials are not required or used as a hiring basis.
Benefits
- Role is located in Europe.
- Compensation is location dependent.
Tech Stack
Categories
DevOpsSite Reliability
About Strike
Strike is the global bitcoin app. It's the simple, fast, and secure way to buy bitcoin and send money globally. Strike is available in 100 countries and territories, including the U.S., Europe, Latin America, and Africa. Sign up in seconds and start with as little as a cent. Join the future of money.