2 months ago
Responsibilities
- Turn ambiguous infrastructure problems into proposals and drive them through RFCs and cross-team architecture reviews.
- Design self-service platform capabilities and APIs, primarily in Go, for onboarding, provisioning, deployment, observability defaults, and day-two operations.
- Set delivery standards using Terraform, GitOps with Argo CD, progressive rollout, testing, and continuous deployment.
- Evolve multi-tenant EKS foundations for reliability, security, scale, and cost, including Envoy Gateway ingress, traffic routing, and multi-region cross-account connectivity.
- Improve SLOs, alerting, incident follow-up, runbooks, and on-call health using Grafana Cloud.
- Develop safe, auditable, human-reviewed AI-assisted operational workflows for alert enrichment, incident context gathering, diagnosis, remediation recommendations, onboarding, and readiness.
- Stay hands-on in the codebase while setting technical direction and driving platform investments through production adoption.
- Join the on-call rotation after onboarding and shadowing and participate in blameless postmortems.
Requirements
- 8+ years of professional, hands-on, full-time software engineering experience in backend, infrastructure, or platform engineering.
- Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
- Strong software engineering skills in Go or a similar language, including design, testing, debugging, review, and maintainability.
- Track record designing, shipping, and operating cloud services or infrastructure platforms in production.
- Deep expertise in at least one of Kubernetes, networking, cloud platforms, reliability engineering, or developer platforms, plus solid Linux, networking, and production-operations fundamentals.
- Experience setting technical direction and leading work requiring cross-team alignment.
- Clear written and verbal communication in a remote environment, including RFCs, design documents, and incident writeups.
- EKS and ingress, CNI, or service-mesh experience is preferred.
- Observability experience with OpenTelemetry, Prometheus, or Grafana is preferred.
- CI/CD and progressive-delivery experience with GitHub Actions, Argo CD, or canaries is preferred.
- Experience leading migrations or adoption programs across teams is preferred.
Benefits
- Remote-first work arrangement with offices in Seattle and Paris
- Flexible work schedule and quarterly Whaleness Days plus an end-of-year Whaleness break
- Home office setup support
- 16 weeks of paid parental leave after six months of employment
- Technology stipend of $100 USD net per month
- PTO plan
- Training stipend for conferences, courses, and classes
- Equity participation
- Medical benefits, retirement, and holidays varying by country
- Docker Swag
About Docker
At Docker, we simplify the lives of developers who are making world-changing apps. Docker helps developers bring their ideas to reality by conquering the complexity of app development. We simplify and accelerate workflows with an integrated development pipeline and application components. Actively used by millions of developers around the world, Docker Desktop and Docker Hub provide unmatched simplicity, agility and choice.
