1 month ago
San Francisco, CA, USA or New York, NY, USAStaff+
Base Salary
$260k - $300k/yr
Responsibilities
- Build integrations that ingest and synchronize customer systems, including EHRs, schedulers, warehouses, and APIs.
- Design transformations and normalized data models for accurate patient and provider profiles.
- Build and optimize streaming and batch data pipelines serving millions of events.
- Power natural-language query interfaces over healthcare data.
- Own data modeling, query performance, data freshness, architecture, implementation, monitoring, reliability, and incident response for production systems.
Requirements
- Experience building and operating production data platforms that ingest, process, and serve millions of events with high reliability.
- Experience designing scalable streaming and batch pipelines, data models, and ETL/ELT workflows.
- Deep expertise in SQL, distributed query optimization, and large-scale data processing.
- Hands-on experience with modern data platforms such as Databricks, Snowflake, Delta Lake, Apache Iceberg, Spark, or Kafka.
- Experience with event-driven architectures, change data capture, online serving systems, reverse ETL pipelines, connector frameworks, or enterprise data integrations.
- Ability to balance performance, scalability, cost, and operational simplicity in distributed systems.
- Preferred experience includes regulated or high-reliability industries, FHIR or HL7, AI or machine-learning data platforms, lakehouse technologies, streaming, CDC, observability, and cost-efficient reliability practices.
- Ability to work onsite in New York City or San Francisco.
Benefits
- Comprehensive health, dental, and vision insurance.
- Daily catered lunch and dinner.
- Mental health support, wellness coaching, and a flexible wellness stipend.
- Annual learning and conference budgets, an annual team offsite, and academic collaboration opportunities.
- Unlimited PTO.
- The team works in person in New York City or San Francisco.
Tech Stack
Categories
Data Engineering
