6 months ago
Base Salary
$165k - $330k/yr
Responsibilities
- Partner with Sales on customer discovery calls
- Lead technical demos and scope architectures, success criteria, and deployment approaches
- Own benchmarking and repeatable deployments across LLMs, embeddings, image and video generation, and VoiceAI use cases
- Advise on GPU selection and latency-versus-throughput tradeoffs
- Use and understand inference runtimes including vLLM, SGLang, and TRT-LMM
- Scope and manage POCs, timelines, deliverables, and next steps
- Coordinate Sales, Engineering, and Forward Deployed Engineering support for customer projects
Requirements
- AI/ML background and the ability to discuss AI/ML topics credibly with technical stakeholders
- Strong customer-facing communication skills, including structured discovery and requirements clarification
- Technical depth to scope solutions without writing production code
- Ability to script and prototype as needed, including comfort with rapid technical experimentation
Benefits
- 100% medical, dental, and vision insurance coverage for employees and dependents
- Flexible paid time off and company-wide Winter Break from Christmas Eve through New Year's Day
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Meaningful equity and exposure to a variety of ML startups
Categories
Solutions Engineering
About Baseten
Inference is everything. Baseten is an AI infrastructure platform giving you the tooling, expertise, and hardware needed to bring great AI products to market - fast. Our proprietary Inference Stack utilizes the cutting-edge of performance research combined with highly performant and reliable infrastructure to give you out-of-the-box global availability with 99.99% of uptime.
