3 hours ago
Remote, EMEASenior
Base Salary
$145k - $290k/yr
Responsibilities
- Partner with Sales on customer discovery calls.
- Lead technical demos and scope architectures, success criteria, and deployment approaches.
- Own benchmarking and repeatable deployments across LLM, embedding, image-generation, video-generation, and Voice AI use cases.
- Advise customers on GPU, latency, throughput, runtime, configuration, and deployment tradeoffs.
- Drive POC execution by defining scope, aligning stakeholders, managing timelines and deliverables, and coordinating with Sales, Engineering, and Forward Deployed Engineering.
- Develop expertise in inference runtimes and common deployment configurations.
Requirements
- AI/ML background and the ability to discuss AI/ML topics credibly with technical stakeholders.
- Strong customer-facing communication skills and the ability to run structured discovery and clarify ambiguous requirements.
- Technical depth to scope solutions without needing to write production code.
- Ability to script and prototype as needed, including comfort with rapid technical experimentation.
Benefits
- Competitive compensation including meaningful equity.
- For U.S. employees, 100% coverage of medical, dental, and vision insurance for employees and dependents.
- Flexible PTO and a company-wide Winter Break from Christmas Eve through New Year's Day.
- Paid parental leave.
- Fertility and family-building stipend through Carrot.
- For U.S. employees, a company-facilitated 401(k).
- Exposure to a variety of ML startups and related learning and networking opportunities.
Categories
Solutions Engineering
About Baseten
Baseten builds an AI inference platform that provides tooling, infrastructure, and hardware to deploy, scale, and serve machine-learning models in production. The company sells managed model serving and developer tooling to software teams at AI product companies, with customers including Notion, Abridge, Writer, and Cursor. Privately held and headquartered in San Francisco, it focuses on high-availability, globally distributed inference for production workloads.
