Sr Site Reliability Engineer
BeyondTrust5 hours ago
Responsibilities
- Design long-term technical solutions and cross-team mechanisms to achieve reliability goals.
- Define and help execute a roadmap for automated, self-service, scalable, observable, and reliable infrastructure services.
- Collaborate with SREs and senior engineers on platform engineering best practices.
- Provide technical guidance and feedback during engineering design reviews for teams onboarding to Platform Infrastructure.
- Automate operational work to reduce toil.
- Build monitoring and alerting for Platform Infrastructure.
- Participate in an on-call rotation and respond to platform availability incidents.
Requirements
- Experience with AWS cloud resources, including S3, EC2, and RDS.
- Experience with Kubernetes clusters running in EKS.
- Experience with the Istio service mesh.
- Experience with infrastructure as code using Terraform or AWS CDK.
- Experience with continuous build using GitLab and continuous delivery tools such as ArgoCD.
- Experience with Datadog.
Benefits
- Flexible culture emphasizing trust and continual learning.
- Diversity and inclusion-focused workplace.
- Employee support programs focused on enabling employees to serve customers effectively.
Tech Stack
Categories
DevOpsSite Reliability
About BeyondTrust
BeyondTrust builds identity and privileged access management software that controls admin rights, manages passwords and keys, and monitors privileged sessions across cloud and on-prem systems. The privately held company, founded in 1985 and headquartered in Johns Creek, Georgia, emerged after Bomgar acquired BeyondTrust and adopted its name. It serves more than 20,000 customers, including 75 of the Fortune 100, via enterprise subscriptions and partner channels.