
Senior Site Reliability Engineer (SRE)
PrizePicksabout 3 hours ago
Remote, Worldwide or Atlanta, GA, USASenior
Base Salary
$120k - $175k/yr
Responsibilities
- Design, implement, maintain, and monitor reliable production systems at scale.
- Lead incident response, mitigate production issues, and conduct post mortem analysis.
- Proactively monitor performance, analyze system failures, identify bottlenecks, and propose solutions.
- Create and support observability/monitoring tools and vendor integrations.
- Drive the growth of a reliability culture, promoting cross-functional collaboration.
- Train and mentor other engineers.
Requirements
- 5+ years of experience as a reliability-focused engineer in a fast-paced environment.
- Deep understanding of cloud computing (AWS, Azure, GCP) and infrastructure as code tools.
- Experience developing applications in languages such as Python, Ruby, or Go.
- Proficient in deploying and supporting applications in Kubernetes at scale.
- Experience with monitoring tools like Grafana, New Relic, or Datadog.
- Familiarity with reliability principles and ability to work cross-functionally.
Benefits
- Company-subsidized medical, dental, and vision plans.
- 401(k) plan with company match.
- Annual bonus and flexible PTO.
- Generous paid leave programs, including 16-week paid parental leave.
- Workplace flexibility and modern work schedules.
- Company-wide in-person events and team outings.
- Lifestyle enhancement program and company equipment provided.
- Annual performance reviews with opportunities for growth.