Base Salary
$150k - $220k/yr
Responsibilities
- Improve core inference services across networking, speech processing, model orchestration, and observability.
- Develop integrations with in-house, third-party, and open-source AI models for perception and conversational dynamics.
- Debug complex system issues involving networking, scheduling, and highly concurrent workloads.
- Rapidly customize backend services to support customer needs.
- Partner with Product to design and implement new services, features, and products end to end.
- Build secure, robust, scalable, testable, and observable services for speech processing and voice agents.
- Actively use and experiment with advanced AI tools as part of the company’s AI-first operating model.
Requirements
- At least 3 years of industry experience.
- Programming experience in Rust or C/C++, with competence in Python.
- High level of experience and understanding of version control, preferably Git.
- Comprehensive experience with UNIX-style systems.
- Excellent written and verbal communication and organizational skills.
- Preferred experience with low-latency, multi-model orchestration for AI-enabled applications.
- Preferred experience with audio processing.
Benefits
- Medical, dental, and vision benefits
- Annual wellness stipend
- Mental health support
- Life, short-term disability, and long-term disability income insurance plans
- Unlimited paid time off
- Generous paid parental leave
- Flexible schedule
- 12 paid US company holidays
- Quarterly personal productivity stipend
- One-time home office upgrade stipend
- 401(k) plan with company match
- Tax savings programs
- Learning and education stipend
- Participation in talks and conferences
- Employee Resource Groups
- AI enablement workshops and sessions
- Benefits are administered locally through an Employer of Record in many countries and vary by region
About Deepgram
Deepgram is the real-time API platform powering the trillion-dollar Voice AI economy. Backed by a $130M Series C at a $1.3B valuation, Deepgram is trusted by 200,000+ developers and 1,300+ organizations to build Voice AI products, platforms, and autonomous agents with the lowest latency, highest accuracy, and enterprise reliability. Our voice-native foundation models and runtime infrastructure have processed 50,000+ years of audio and over 1 trillion words, making Deepgram the most experienced voice AI platform in the world. Industry-leading models & platform: 👂 Nova-3 — the world’s most accurate real-time speech-to-text model 🔊 Aura-2 — professional, enterprise-grade text-to-speech 💬 Flux — the first Conversational Speech Recognition model designed to handle interruptions 🚀 Voice Agent API — enterprise-ready, real-time conversational AI 🧠 Saga — the Voice OS Beyond core infrastructure, Deepgram is expanding the Voice AI ecosystem through: 💪 Powered by Deepgram, supporting voice products built by leading AI startups and enterprise organizations 🌉 A new Voice AI Collaboration Hub in San Francisco for builders, partners, and the voice community 🍔 The acquisition of OfOne, delivering real-time Voice AI for restaurants and drive-thru operations with 95%+ containment 📃 A growing patent portfolio in Voice AI Much like APIs powered the payments and cloud economies, Deepgram is building the foundation for a trillion-dollar B2B Voice AI economy—centered on the most natural human interface: voice.
