over 2 years ago
San Francisco, CA, USA or New York, NY, USAMid Level
H1B sponsor
Base Salary
$165k - $330k/yr
Responsibilities
- Develop and maintain production software systems and product features using general-purpose programming languages, preferably Python.
- Design, implement, deploy, and monitor Baseten solutions end to end in partnership with customer engineering teams.
- Translate vague objectives into clear specifications and well-defined proofs of concept, then rapidly ship well-tested services.
- Optimize AI/ML projects and contribute to improvements across the technical stack.
- Develop features and product requirements documents with engineering and product teams.
- Own customer projects and products from problem framing through production deployment, acting across engineering, project management, and product management.
- Make effective tradeoffs under ambiguity and avoid unnecessary complexity.
Requirements
- Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or a related field.
- At least 2 years of professional work experience in a fast-paced, high-growth environment.
- Experience with one or more general-purpose programming languages in a production-level environment, with a strong preference for Python.
- Familiarity with AI/ML pipelines and the lifecycle of ML model development and deployment.
- Strong communication skills, particularly when discussing complex technical topics.
- Experience building or optimizing AI/ML projects is highly valued.
Benefits
- 100% coverage of medical, dental, and vision insurance for employees and dependents.
- Flexible PTO, including a company-wide Winter Break with offices closed from Christmas Eve through New Year's Day.
- Paid parental leave.
- Fertility and family-building stipend through Carrot.
- Company-facilitated 401(k).
- Exposure to a variety of ML startups and related learning and networking opportunities.
About Baseten
Baseten builds an AI inference platform that provides tooling, infrastructure, and hardware to deploy, scale, and serve machine-learning models in production. The company sells managed model serving and developer tooling to software teams at AI product companies, with customers including Notion, Abridge, Writer, and Cursor. Privately held and headquartered in San Francisco, it focuses on high-availability, globally distributed inference for production workloads.
