1 day ago
Base Salary
$201k - $332k/yr
Responsibilities
- Own the long-term architecture and technical direction for Nordstrom’s cloud platform, CI/CD, and SRE capabilities.
- Establish design principles, reliability targets, degradation strategies, and standards enforced through pipelines and production validation.
- Advance autonomous deployments, predictive runtime scaling, multi-region resilience, and platform efficiency.
- Drive SRE practices including SLOs, error budgets, observability standards, incident command, reliability tooling, load testing, and chaos exercises.
- Guide significant incidents across systems and convert incident learnings into platform defaults and organizational improvements.
- Influence Platform, SRE, Security, product, and engineering teams without direct authority while resolving cross-team technical disagreements.
- Mentor principal and senior engineers, guide technical design reviews and hiring, and represent Nordstrom’s platform engineering externally.
Requirements
- 15+ years of professional software or infrastructure engineering experience.
- Bachelor’s or Master’s degree in Computer Science, Engineering, or a related field.
- Deep hands-on architectural experience with large-scale cloud platforms, production Kubernetes workloads, multi-account cloud estates, and centrally owned shared services.
- Enterprise-scale continuous delivery experience including progressive delivery, canary analysis, immutable artifacts, versioned configuration, promotion-over-rebuild practices, and rollback design.
- Experience with software supply chain security including build provenance, artifact signing or attestation, and deployment-time policy enforcement.
- Experience establishing SRE practices at scale, including SLOs, error budgets, observability, incident command, and reliability measurement.
- Architectural experience with cell-based isolation, multi-region active-passive and active/active approaches, concurrent-write consistency, and rehearsed failover.
- Experience engineering for predictable seasonal traffic peaks, including capacity forecasting, load and scale testing, and chaos exercises.
- Sound financial judgment regarding reliability investments and experience influencing organizations without formal authority.
- Informed judgment about AI for operations and delivery, organizational design, reliability ownership boundaries, and centralized operations failure modes.
- Strong written and verbal communication skills for technical and executive audiences.
Benefits
- Medical, vision, dental, retirement, paid time away, life insurance, and disability benefits.
- Merchandise discount and Employee Assistance Program resources.
- 401(k), paid time off accruals, holidays, and additional benefits subject to eligibility requirements.
- The role is hybrid in Seattle, Washington, and requires working in the Nordstrom corporate headquarters at least four days per week.
- The position may be eligible for performance-based incentives or bonuses.
Tech Stack
Categories
DevOpsSite Reliability
About Nordstrom
Nordstrom is a U.S. fashion retailer selling apparel, shoes, accessories, beauty, and home goods through department stores, off-price Nordstrom Rack, and e-commerce. Founded in 1901 and headquartered in Seattle, it is a public company (NYSE: JWN) serving consumers nationwide, with technology, merchandising, and supply-chain teams supporting its digital and in-store operations.
