about 3 hours ago
Base Salary
$296k - $346k/yr
Responsibilities
- Build and operate core cloud platform services for compute lifecycle and maintenance workflows.
- Design reliable APIs, backend services, and orchestration systems for Lambda’s GPU cloud.
- Work on bare metal lifecycle systems including launch, terminate, and host reclaim workflows.
- Improve deployment, observability, testing, and operational readiness for control-plane services.
- Debug complex production issues across distributed services and infrastructure dependencies.
- Collaborate with cross-functional teams to define contracts and deliver cloud capabilities.
- Contribute to architecture, design docs, code reviews, and mentoring within the team.
Requirements
- Bachelor's degree or equivalent working experience.
- 6+ years of professional software engineering experience in backend or distributed systems.
- Strong proficiency in Python, Go, or a similar backend/system language.
- Experience designing and operating APIs, workflow engines, and orchestration services.
- Understanding of reliability fundamentals such as fault tolerance and production debugging.
- Familiarity with cloud infrastructure primitives like compute, networking, and storage.
- Comfortable with Linux, containers, Kubernetes, and infrastructure automation.
- Experience owning production services and participating in on-call duties.
Benefits
- Generous cash and equity compensation.
- Health, dental, and vision coverage for employees and dependents.
- Wellness and commuter stipends for select roles.
- 401k Plan with 2% company match for USA employees.
- Flexible paid time off plan.
