1 day ago
Base Salary
$144k - $194k/yr
Responsibilities
- Design, build, and operate offline-first fleet management subsystems supporting provisioning, state reconciliation, staged rollouts, and atomic rollback across multi-processor machines.
- Develop cloud-side services that orchestrate, monitor, and coordinate fleets of edge devices across disconnected or intermittently connected environments.
- Build secure OTA delivery pipelines for AI models and software across heterogeneous processor topologies.
- Write and review high-performance production code and debug issues spanning operating systems, runtimes, networks, and physical devices.
- Own subsystems end to end from design documentation and architectural decisions through implementation, testing, and production operation.
- Define correctness guarantees for disconnected, reconnecting, and partially connected system states.
- Participate in on-call rotation, troubleshoot production issues, and drive root-cause fixes.
- Mentor junior engineers and contribute through code reviews, design reviews, technical guidance, roadmap planning, and architectural direction.
- Collaborate with product, hardware, and adjacent platform teams on requirements and integration points.
Requirements
- At least 3 years of non-internship professional software development experience.
- Bachelor's degree or equivalent.
- Knowledge of professional software engineering practices across the full software development life cycle, including coding standards, software architecture, code reviews, source control management, continuous deployments, testing, and operational excellence.
- Strong Linux and systems-level skills, including production debugging across the hardware-software boundary.
- Practical distributed-systems experience with state machines, idempotency, reconciliation, retries, and persistence.
- Ability to design offline-first and reconnect-and-reconcile behavior for edge systems without guaranteed connectivity.
- Experience developing cloud-side services that orchestrate, monitor, and coordinate with fleets of edge devices.
- Track record of independently owning substantial subsystems from design through production and authoring sound design documents and architectural decisions.
- Experience with fleet management or OTA update systems, including staged rollout, rollback, pause-resume, and dependency handling.
- Preferred experience includes edge containers such as containerd, K3s, and KubeEdge; heterogeneous x86, ARM, GPU, and NPU platforms; accelerator or model versioning; ROS 2, DDS middleware, NVIDIA Jetson, or Isaac; TUF, Uptane, TPM, HSM, and secure boot; reliability-critical domains; artifact or model distribution; and AWS IoT or AWS IoT Greengrass.
Benefits
- Comprehensive health insurance including medical, dental, vision, prescription, Basic Life and AD&D insurance, and optional supplemental life plans.
- Employee assistance, mental health support, medical advice line, flexible spending accounts, and adoption and surrogacy reimbursement coverage.
- 401(k) matching, paid time off, and parental leave.
- Flexible work hours and arrangements, employee-led affinity groups, learning and development opportunities, mentorship, and career-growth resources.
- The compensation package may include sign-on payments and restricted stock units in addition to base salary.
- The position is based in Seattle, Washington.
Tech Stack
About Amazon
Amazon builds and operates a global e-commerce marketplace, logistics network, and consumer devices, and provides cloud computing via AWS for businesses and developers. The company earns revenue from online retail, third‑party seller services, subscriptions like Prime, advertising, and AWS usage. Founded in 1994 and headquartered in Seattle, it is publicly traded on NASDAQ (AMZN) and serves customers in dozens of countries.
