1 month ago
Base Salary
$140k - $280k/yr
Responsibilities
- Own core infrastructure including databases, AWS services, networking, and CI/CD.
- Build systems that support thousands of concurrent AI operations with sub-second response times and cost efficiency under load.
- Design observability for system errors and business-level silent failures.
- Orchestrate parallel agentic workloads reliably and cost-effectively as scale increases.
- Build tooling and abstractions that improve product engineering velocity.
- Own reliability through SLOs, on-call support, incident response, and post-mortems.
- Build evaluation infrastructure to measure improvements in AI outputs.
Requirements
- Software engineering experience focused on platform, infrastructure, or SRE work, with level determined during interviews.
- Proficiency in Python, TypeScript, Go, or a similar language.
- Production experience with AWS, Postgres, and observability tooling.
- Experience building systems and tooling relied upon by other engineers.
- Experience owning production systems at scale and responding when they fail.
- Based in San Francisco or willing to relocate.
- Preferred qualifications include experience orchestrating AI/ML workloads, agent frameworks, LLM inference infrastructure, voice AI, real-time systems, deep AWS experience with ECS, Lambda, RDS, and VPC design, and startup experience in hypergrowth companies.
Benefits
- Salary of $140,000–$280,000 depending on experience, plus performance bonuses and equity.
- San Francisco, in-office role; candidates must be based in San Francisco or willing to relocate.
- Monday–Friday schedule with a very early morning start and five days per week in the office.
- Uber commuter benefits.
- Breakfast, lunch, dinner, snacks, drinks, and coffee provided daily.
- Free gym membership.
- Health, dental, and vision insurance.
