2 months ago
Toronto, Canada or London, United KingdomSenior / Staff+
Responsibilities
- Design, build, and own CI pipelines and test execution across pull request, merge, and nightly workflows for the DX-1 software stack.
- Scale test execution using staged lanes, parallelism, real test isolation, and content-addressed caching.
- Manage heterogeneous cloud and self-hosted CI fleets, including virtual machines, containers, bare-metal systems, and scarce hardware resources.
- Provide fair, monitored, fail-fast shared access to simulation compute, emulator systems, accelerator boards, prototype platforms, and other hardware-in-the-loop resources.
- Build deterministic performance regression baselines, metric aggregation, CI health dashboards, and product readiness dashboards.
- Select and operate metrics storage for long-lived, high-cardinality time-series data and observable software.
- Define hermetic, reproducible, least-privilege, fail-closed build and test standards while containing misconfiguration and untrusted-job blast radius.
- Partner with infrastructure, compiler, runtime, simulator, and modelling teams and influence cross-functional standards without formal authority.
Requirements
- Experience owning large-scale build/test infrastructure, CI/CD, developer productivity, or release engineering systems end to end.
- Experience scaling large test suites with staged lanes, isolated parallelism, content-addressed caching, and cost-effective feedback.
- Experience managing heterogeneous CI runner fleets across cloud and on-premises environments, including virtual machines, containers, and bare-metal hardware.
- Experience providing reservations, remote access, hardware-in-the-loop testing, and portable artifacts for scarce or expensive resources.
- Experience with performance regression baselines, deterministic testing, noise handling, metric aggregation, attribution, and bisection.
- Experience building observability and metrics platforms, CI health dashboards, product readiness dashboards, and long-lived high-cardinality time-series stores.
- Strong scripting and systems programming skills, including Python and a systems language, plus proficiency with containers, Linux, and cloud infrastructure such as AWS.
- Knowledge of hermetic builds, least privilege, fail-closed defaults, and blast-radius control.
- Excellent communication and cross-functional influencing skills.
- Bachelor’s degree or higher in Computer Science, Electrical Engineering, Mathematics, or a related field.
- Preferred experience with GitHub Actions or comparable CI at scale, AWS-based CI runner fleets, hardware-in-the-loop or lab automation, custom silicon or FPGA bring-up, Prometheus, Grafana, Datadog, columnar warehouses, HPC, cluster batch scheduling, release engineering, or developer-productivity platforms.
Benefits
- Meaningful stock options and ownership.
- Employer-contributed retirement plans.
- Annual Living-Local Bonus for employees residing within 20 minutes of the office.
- The posting states that work eligibility is subject to U.S. export-control citizenship or permanent-residency restrictions for certain countries.
