
Baseten
Baseten builds an AI inference platform that provides tooling, infrastructure, and hardware to deploy, scale, and serve machine-learning models in production. The company sells managed model serving and developer tooling to software teams at AI product companies, with customers including Notion, Abridge, Writer, and Cursor. Privately held and headquartered in San Francisco, it focuses on high-availability, globally distributed inference for production workloads.
Open Positions at Baseten
31 open positions
Build and ship AI-powered workflows, agents, dashboards, and automations that improve GPU capacity planning and operations at Baseten. This hands-on role combines AI application development, workflow automation, systems integration, and operational problem-solving.
Lead the global GPU capacity function at Baseten, combining infrastructure engineering, Kubernetes orchestration, and financial modeling to keep AI workloads reliable and cost-efficient across clouds. You’ll manage specialized GPU pods and build systems that scale capacity operations by 10x.
Own the identity and authorization layer of Baseten’s AI platform, building fine-grained access controls, credential systems, and enterprise administration experiences. This founding role combines backend systems engineering, security-focused architecture, and technical leadership.
Build AI-powered workflows, agents, dashboards, and integrations that make Baseten’s go-to-market organization faster and more effective. This hands-on GTM Engineering role combines automation, RevOps systems ownership, and custom internal application development.
Build and ship agentic AI product experiences, internal automations, and reliable execution systems for customers training and post-training frontier models. This hands-on AI Engineer role spans research collaboration, backend and frontend development, evaluation, and production delivery.
Forward Deployed Engineers own the technical success of major AI customers, taking ambiguous problems from discovery through production while optimizing inference, post-training, and deployment systems. This customer-facing role combines hands-on software engineering, incident response, and product influence across mission-critical AI workloads.
Build Baseten's internal AI developer platform, from agent context infrastructure and MCP servers to evaluation, rollout, and safety tooling. You'll help engineering teams adopt AI through useful defaults and self-service infrastructure rather than mandates.
Build the testing frameworks and infrastructure that help Baseten engineers ship reliable, mission-critical AI inference systems. You will own testing strategy, resilience and performance tooling, ephemeral environments, and feedback loops used across the organization.
Build the deployment platform that helps Baseten engineers ship AI products quickly and safely across Kubernetes clusters worldwide. You’ll own continuous delivery, progressive rollouts, automated rollback, and foundational developer infrastructure.
Build the observability foundations behind Baseten’s AI infrastructure, including high-throughput telemetry pipelines, storage, instrumentation, alerting, and SLO systems. As an early member of the team, you’ll shape how internal and external customers detect and resolve operational issues at scale.
Unlock 21 More Matching Jobs
Create an account to view all jobs matching your criteria.
First to know
Discover the latest jobs before everyone else
Zero Spam
No ghost jobs, reposts, or sponsored listings
AI-powered filters
Find the most relevant jobs for you