3 months ago
Berlin, Germany +2 moreSenior
Responsibilities
- Develop the Managed PostgreSQL control plane and lifecycle automation for provisioning, high availability, failover, backups, point-in-time recovery, version upgrades, and zero-downtime maintenance.
- Tune and harden PostgreSQL internals and turn replication, WAL, vacuum, query planning, connection pooling, and extension capabilities into product features and customer-friendly defaults.
- Build migration tooling for AWS RDS, Google Cloud SQL, Azure Database for PostgreSQL, and self-managed clusters with minimal downtime.
- Develop the AI-Postgres experience, including vector search with pgvector and pgvectorscale, hybrid retrieval, and integration with the Nebius AI Cloud stack.
- Operate the service with SRE practices by defining SLOs, building observability, leading incident response, and incorporating postmortem findings into the platform.
- Work directly with customers on architecture reviews, performance escalations, and complex production issues.
Requirements
- At least five years of professional software engineering experience, including significant experience building or operating production PostgreSQL at scale.
- Strong software engineering skills in Go or another backend or systems language, with willingness to work primarily in Go.
- Deep knowledge of PostgreSQL internals, including MVCC, WAL, physical and logical replication, vacuum, query planning, extensions, and partitioning.
- Hands-on experience with Patroni, Stolon, pg_auto_failover, pgBackRest, WAL-G, pgBouncer, PgCat, and logical replication tooling.
- Production database expertise, including EXPLAIN ANALYZE, lock behavior, database corruption, replication topologies, and large-scale datasets.
- Ability to write reliable code and investigate complex problems.
- Teamwork-oriented approach.
- Experience with pgvector and pgvectorscale, vector search at scale, index selection, and recall-versus-latency trade-offs is preferred.
- Experience building managed-database control planes at a cloud provider is preferred.
- Experience writing Kubernetes operators with Go, controller-runtime, or kubebuilder is preferred.
- Contributions to PostgreSQL, popular extensions, or the surrounding open-source ecosystem are preferred.
- AI/ML workload experience, including RAG pipelines, embedding stores, or GPU-resident workloads, is preferred.
Benefits
- Competitive compensation
- Career growth and learning opportunities
- Flexibility and work-life balance
- Collaborative and innovative culture
- Opportunity to work on impactful AI projects
- International environment and talented teams
- Hybrid or remote work from EU time zones, with office options in Amsterdam and London
Tech Stack
Categories
BackendSite Reliability
About Nebius
Nebius builds a full-stack AI cloud offering GPU compute, storage, and tools for training and deploying ML models for startups, enterprises, and research labs. It sells consumption-based cloud infrastructure (IaaS/PaaS) and managed services tailored to generative AI workloads, including large-scale model training and inference. The company is headquartered in Amsterdam and operates as an independent provider.
