2 months ago
Remote, WorldwideSenior
Responsibilities
- Own the reliability of deployment and release systems and their control plane against SLOs and error budgets.
- Standardize and instrument fragmented deployment workflows so pre-production provides trustworthy signals.
- Drive disaster-recovery readiness and reproducible deployments from scratch.
- Build and operate health and SLO monitoring for critical user flows using synthetic testing.
- Reduce mean time to detect and recover from deployment-related incidents.
- Participate in on-call, lead blameless postmortems, and convert findings into runbooks, alerting, and automation.
- Improve deployment observability and auditability by recording what shipped, where, when, and by whom.
- Document break-glass paths, access models, and operational runbooks.
- Define and track SLAs, SLOs, error budgets, and DORA delivery metrics with meaningful alerting.
- Ensure deployments fail fast and safely when health checks degrade.
- Harden access and break-glass workflows for safe incident response.
- Partner with product engineering and platform teams to align release practices with reliability and availability targets.
Requirements
- At least 5 years of experience in SRE, production operations, platform engineering, or release engineering.
- Experience operating production systems at scale and participating in on-call rotations.
- Fluency with SLAs, SLOs, error budgets, DORA metrics, operational KPIs, and observability tooling.
- Experience leading incident response with incident.io, PagerDuty, Opsgenie, or similar tools.
- Production experience with AWS, including multiple accounts, IAM, and VPC.
- Comfort with infrastructure-as-code using Pulumi or Terraform and with Kubernetes.
- Ability to script and automate to eliminate operational toil.
- Clear communication with infrastructure specialists and product engineers.
- Ability to work effectively in asynchronous, globally distributed teams and navigate ambiguity.
Benefits
- Fully remote with no Supabase offices and a WeWork membership or coworking allowance worldwide.
- ESOP equity ownership for every team member.
- Tech allowance for equipment such as a laptop, monitor, or headphones.
- Health insurance covered at 100% for employees and 80% for dependents.
- Annual company off-site in a new city for one week.
- Flexible asynchronous work environment.
- Annual education allowance for courses, books, conferences, or other professional development.
Tech Stack
Categories
DevOpsSite Reliability
About Supabase
Supabase builds an open-source Postgres-based backend platform for developers, bundling database, authentication, storage, realtime APIs, edge functions, and vector search. It monetizes through a managed cloud service and enterprise offerings while remaining deployable self-hosted. Founded in 2020 and privately held, the company operates globally with a fully remote team. Developers use Supabase as a Firebase alternative to launch quickly and scale production applications on PostgreSQL.
