3 months ago
Responsibilities
- Identify architectural changes that improve reliability, performance, and availability.
- Foster a culture of reliability across the engineering organization.
- Design and implement deployment, upgrade, rollback, and postmortem review processes.
- Participate in an on-call rotation and respond to production incidents.
- Build monitoring systems that ensure high-quality service for customers.
- Debug production issues across all services and levels of the stack.
- Own the full development lifecycle, including building, shipping, supporting, maintaining infrastructure, and measuring impact.
Requirements
- A record of exercising full ownership and agency to improve product reliability and uptime.
- Ability to learn new frameworks quickly and work across the backend and frontend stack.
- Ability to participate in an on-call rotation and respond to production incidents.
- Ability to work in person in Speakeasy’s San Francisco office.
Benefits
- In-person work in the San Francisco office.
- Participation in an on-call rotation is part of the role.
Categories
DevOpsSite Reliability
About Speakeasy
The enterprise control plane for AI. Secure and centrally manage MCPs, Skills, and Assistants your whole company can use, with fine-grained permissions, threat detection, and full observability.
