7 months ago
Base Salary
$165k - $330k/yr
Responsibilities
- Partner with Sales on customer discovery calls
- Lead technical demos and scope architectures, success criteria, and deployment approaches
- Own benchmarking and repeatable deployments across LLMs, embeddings, image and video generation, and VoiceAI use cases
- Advise on GPU selection and latency-versus-throughput tradeoffs
- Use and understand inference runtimes including vLLM, SGLang, and TRT-LMM
- Scope and manage POCs, timelines, deliverables, and next steps
- Coordinate Sales, Engineering, and Forward Deployed Engineering support for customer projects
Requirements
- AI/ML background and the ability to discuss AI/ML topics credibly with technical stakeholders
- Strong customer-facing communication skills, including structured discovery and requirements clarification
- Technical depth to scope solutions without writing production code
- Ability to script and prototype as needed, including comfort with rapid technical experimentation
Benefits
- 100% medical, dental, and vision insurance coverage for employees and dependents
- Flexible paid time off and company-wide Winter Break from Christmas Eve through New Year's Day
- Paid parental leave
- Fertility and family-building stipend through Carrot
- Company-facilitated 401(k)
- Meaningful equity and exposure to a variety of ML startups
Categories
Solutions Engineering
About Baseten
Baseten builds an AI inference platform that provides tooling, infrastructure, and hardware to deploy, scale, and serve machine-learning models in production. The company sells managed model serving and developer tooling to software teams at AI product companies, with customers including Notion, Abridge, Writer, and Cursor. Privately held and headquartered in San Francisco, it focuses on high-availability, globally distributed inference for production workloads.
