19 hours ago
Base Salary
$210k - $275k/yr
Responsibilities
- Design, build, and operate backend systems powering Ambient AI products across mobile and web.
- Build inference workflows coordinating transcription, diarization, language models, clinical extraction, summarization, and other AI capabilities.
- Develop durable orchestration for scheduling, queueing, retries, timeouts, fallbacks, replay, idempotency, model routing, and workflow recovery.
- Improve service reliability, scalability, observability, deployment systems, capacity planning, and operational tooling.
- Define SLOs and lead incident response and post-incident analysis for critical production workflows.
- Design APIs and data models for long-running workflows and build tools for pipeline inspection, testing, replay, evaluation, and debugging.
- Partner with mobile, product, AI, security, and clinical teams and mentor engineers on distributed systems and backend architecture.
Requirements
- 8+ years of professional backend or infrastructure engineering experience.
- Experience designing, building, and operating production distributed systems.
- Strong proficiency in at least one modern backend programming language.
- Experience with asynchronous processing, message queues, event-driven systems, or durable workflow execution.
- Strong understanding of idempotency, consistency, concurrency, backpressure, retries, failure isolation, and eventual completion.
- Experience operating cloud services, including deployment, monitoring, scaling, and incident response.
- Experience designing APIs, service boundaries, and data models for complex product workflows.
- Track record of improving reliability, observability, scalability, or operational efficiency.
- Strong debugging skills across application, infrastructure, data, and external dependency boundaries.
- Ability to reason about real-time and long-running workloads with different latency and durability requirements.
- Nice-to-have experience with speech-to-text, diarization, audio ML, LLM or multi-model inference pipelines, Temporal, Kafka, Pub/Sub, SQS, model routing, inference gateways, rate limiting, batching, caching, GPU-backed workloads, Grafana, OpenTelemetry, Sentry, containers, Kubernetes, infrastructure as code, cloud-native deployment, offline clients, regulated environments, or AI-native products.
Benefits
- Build backend foundations for an AI healthcare product used in clinical workflows.
- Work on distributed orchestration and production AI systems with direct clinician impact.
- Maintain deep ownership of backend architecture while partnering with product and AI teams.
- Help shape technical strategy, operational standards, and engineering culture at a growing platform.
- Opportunity to grow into broader backend, infrastructure, or technical leadership.
Tech Stack
Categories
About Commure
Commure builds an AI-native healthcare platform for providers and health systems, spanning ambient clinical documentation, provider copilots, coding, and revenue cycle automation, integrated with 60+ EHRs. It sells enterprise software and services to hospitals and care organizations to streamline clinical and administrative workflows. The privately held company is headquartered in Mountain View, California, and reports adoption by 500,000+ clinicians across 500+ organizations, with billions in claims processed annually.
