4 days ago
Remote, United StatesSenior
Base Salary
$125k - $135k/yr
Responsibilities
- Own monitoring and alerting for MySQL InnoDB Clusters and PostgreSQL across multiple datacenters.
- Establish and execute quarterly database backup verification and disaster recovery testing.
- Respond to database incidents through on-call rotations, documented runbooks, escalation, post-incident review, and remediation.
- Monitor replication health and manage database user provisioning, deprovisioning, access audits, and role-based access control.
- Execute security compliance remediation, including penetration-test findings, GRC audit requirements, and encryption-at-rest verification.
- Maintain the Debezium and Kafka Connect data pipeline in coordination with the Kafka infrastructure team.
- Build Puppet profiles and PHP, Python, or Go automation to reduce operational toil.
- Develop runbooks, operational documentation, and disaster recovery procedures while partnering with the Senior Platform Engineer on shared database reliability ownership.
Requirements
- 7+ years of production experience in Site Reliability Engineering, DevOps, or Database Operations at scale.
- Deep production expertise with MySQL, including InnoDB Cluster, Group Replication, MySQL Router, and ProxySQL.
- Production experience with PostgreSQL replication, high availability, performance tuning, and operational management.
- Strong proficiency with configuration management tools, preferably Puppet, and infrastructure-as-code practices.
- Experience with xtrabackup, pg_dump, mysqldump, and database disaster recovery procedures.
- Proficiency in PHP and Python or Go for automation, tooling, and integrations.
- Experience with Prometheus exporters, Grafana dashboards, alerting frameworks, and SLO/error-budget methodology.
- Familiarity with Kafka Connect, Debezium, or similar change-data-capture pipelines.
- Strong incident response and on-call experience, including post-incident review and remediation.
- Excellent communication and cross-team collaboration skills as a senior peer.
Benefits
- Company-paid medical, dental, and vision insurance premiums.
- 401(k) plan matching 100% up to 4% with immediate vesting.
- $2,500 annual professional development reimbursement.
- 11 holidays, paid time off accrual, and PTO rollover.
- Increased PTO at three- and ten-year anniversaries, a one-month paid sabbatical every five years, and an annual anniversary bonus.
- $500 remote office setup stipend in the first year and $400 in each following year.
- Internet reimbursement up to $75 per month and gym membership reimbursement up to $50 per month.
- Company-paid Wellable subscription.
- Applications are currently accepted from candidates residing in the listed U.S. states.
Tech Stack
Categories
Site Reliability
About Vultr
Vultr builds and operates a self-service cloud infrastructure platform offering Cloud Compute, Cloud GPU, Bare Metal, object storage, and managed Kubernetes for developers, startups, and enterprises. It runs in 33 data center locations worldwide and sells on-demand IaaS and GPU capacity, including AI infrastructure. Founded in 2014 and headquartered in West Palm Beach, Florida, Vultr is privately held and announced a private equity financing in 2024.
