11 days ago
Remote, WorldwideMid Level
Responsibilities
- Develop and own solutions that transform raw audio and sparse or missing metadata into organized, ready-to-sell datasets.
- Build prototypes and evolve them into reliable production systems that others can operate.
- Characterize large and unfamiliar audio datasets and assess their suitability for customer use cases.
- Own customer engagements from feasibility through implementation, delivery, post-delivery support, and iteration.
- Translate customer model-development goals into technical plans with clear acceptance criteria.
- Partner with the audio GM, product and engineering teams, Data Lab, and commercial stakeholders.
- Identify recurring customer needs and help design and build reusable cross-vertical platform capabilities.
- Create audio-vertical technical playbooks, quality standards, reusable tooling, and platform investment roadmaps.
- Manage multiple concurrent deals and customer priorities, including availability outside standard hours when deals are live.
Requirements
- 3+ years of experience as an engineer, including meaningful exposure to customers or external technical stakeholders.
- Experience working directly with media data; direct audio-data experience is a plus.
- Experience building and operating systems that process, analyze, or deliver data at scale.
- Ability to translate ambiguous requirements, communicate trade-offs, and build trust with technical stakeholders.
- Demonstrated end-to-end ownership from problem definition through implementation, validation, and support.
- Tolerance for ambiguity, bias to action, and judgment about when to investigate further or push back.
- Comfort working in a fast-moving environment with multiple concurrent priorities and time-sensitive customer work.
- Early-stage company, founding or early engineering, or startup-like broad-ownership experience is preferred.
- Hands-on Python experience is preferred.
- Product engineering experience or a strong product mindset developed in close partnership with users is preferred.
- Experience with search, vector embeddings, semantic retrieval, or ML-assisted data curation is preferred.
- Experience with AWS, Dagster, Databricks, or Vercel is preferred.
- Audio processing, speech or music ML, transcription, diarization, annotation, quality measurement, privacy, or redaction experience is preferred.
