2 months ago
Responsibilities
- Own the technical vision and roadmap for cloud infrastructure, reliability, and platform engineering.
- Design, build, and maintain highly available, scalable, and secure AWS infrastructure.
- Architect and standardize Terraform-based Infrastructure as Code across environments.
- Design and optimize CI/CD pipelines for reliable software delivery.
- Define and implement SLOs, SLIs, error budgets, reliability standards, and incident-management practices.
- Lead observability, monitoring, logging, alerting, incident response, and root-cause analysis initiatives.
- Build self-service platform capabilities and automation for engineering teams.
- Drive containerization, orchestration, infrastructure modernization, and platform scalability.
- Partner with Security and Engineering on cloud security, compliance, application reliability, and operational excellence.
- Lead infrastructure architecture decisions, evaluate technologies, mentor engineers, and influence technical direction across teams.
Requirements
- 10+ years of experience in Site Reliability Engineering, Platform Engineering, DevOps, Cloud Infrastructure, or Software Engineering.
- Proven experience designing and operating large-scale, highly available cloud infrastructure in AWS.
- Strong software engineering background with the ability to write production-quality code and automation.
- Expert-level Infrastructure as Code experience, preferably with Terraform.
- Deep experience designing and maintaining CI/CD pipelines using GitLab or similar platforms.
- Strong knowledge of Kubernetes, containerized workloads, and cloud-native architectures.
- Extensive experience with observability, distributed tracing, logging, monitoring, and incident response.
- Experience defining and implementing SLOs, SLIs, and reliability engineering practices.
- Strong understanding of networking, security, Linux systems administration, and cloud architecture.
- Experience supporting high-traffic SaaS applications and mission-critical production environments.
- Ability to influence technical direction without direct authority and mentor engineers across multiple teams.
- Experience working in Agile development environments and collaborating with cross-functional engineering teams.
Benefits
- 22 days of PTO plus public holidays.
- 401(k) match.
- Medical, dental, and vision insurance.
- Maternity and paternity leave.
- Full-time hybrid position working from the Austin, Texas office in the Arboretum Area.
- Opportunities for career progression and collaborative learning in a fast-paced engineering culture.
Tech Stack
Categories
DevOpsSite Reliability
About ShipperHQ
ShipperHQ provides a SaaS platform that lets eCommerce merchants manage shipping at checkout, including real-time carrier rates, delivery options, and rules-based fulfillment. Headquartered in Austin, Texas and privately held, it supports thousands of brands in 150+ countries across DTC and enterprise retail. The business is subscription-based, offering checkout optimization and shipping rate management tools to control costs and delivery transparency at checkout.
