
Senior Site Reliability Engineer
Honeycomb.io17 days ago
Remote, IrelandSenior
Responsibilities
- Scale backend systems to support high-volume customers.
- Improve infrastructure reliability, scalability, cost efficiency, and developer experience.
- Work with backend teams to optimize the engineering stack and infrastructure.
- Become and later train others as an Incident Commander.
- Participate in the EU side of a follow-the-sun on-call rotation.
- Help navigate tradeoffs between reliability and other organizational priorities.
- Support a healthy cross-Atlantic engineering culture through transparent communication and feedback.
- Optionally represent Honeycomb through blog posts, conference talks, and presentations.
Requirements
- Strong experience with AWS and Kubernetes.
- Experience performing infrastructure cost analysis and reduction.
- Solid experience with Helm, Terraform, and CI/CD.
- Project management skills and software engineering experience.
- Experience with Kafka or another high-volume distributed system.
- Golang and performance engineering experience are preferred.
- Familiarity with observability concepts such as SLOs and instrumentation.
- Strong written and spoken communication skills, including adapting communication to different audiences and giving direct feedback.
- Ability to work through ambiguity, experiment, and collaborate with geographically distributed teams.
- Interest in both the technical and human aspects of reliability engineering.
Benefits
- Fully distributed, remote-first work arrangement.
- Base salary range of €140.590—€165.400 EUR based on experience.
- Generous equity and an employee-friendly stock program.
- Unlimited PTO.
- Home office, co-working, and internet stipend.
- Full benefits coverage for employees, with additional dependent coverage available.
- Up to 16 weeks of paid parental leave.
- Annual development allowance.
- Honeycomb cannot currently sponsor or support visa transfers.
Tech Stack
Categories
DevOpsSite Reliability