about 3 hours ago
Responsibilities
- Operate production day-to-day, including on-call duties and incident response.
- Define and refine SLIs/SLOs and error budgets for reliability practices.
- Strengthen observability across metrics, logs, traces, and alerting.
- Ship infrastructure through code in a GitOps workflow.
- Manage PostgreSQL performance tuning, schema reviews, and online migrations.
- Mentor engineers on reliability and database fundamentals.
Requirements
- 4+ years in SRE, DevOps, Platform/Infrastructure, or backend engineering.
- Hands-on experience with Kubernetes and GitOps workflows.
- Solid working knowledge of PostgreSQL in production environments.
- Understanding of cloud networking fundamentals.
- Proficient with Linux at the operator level.
- Experience in incident response and structured debugging.
- Working proficiency in Go or Python, with strong communication skills.
Benefits
- Competitive Salary & Stock Options.
- Health Benefits.
- One-time USD $500 for New Hire Home-Office Setup.
- Monthly Stipend of USD $150 via a Brex Card.
