Modular

Senior AI Graph Compiler Engineer

Modular
Apply
4 months ago
Remote, United StatesSenior

Base Salary

$198k - $286k/yr

Responsibilities

  • Design and develop compiler optimizations for ML inference efficiency across CPUs, GPUs, and ML accelerators.
  • Compile large-scale workloads onto heterogeneous hardware by mapping hardware-agnostic graphs to hardware-specific device and node graphs.
  • Collaborate with kernels, serving, and Mojo compiler teams on end-to-end performance technologies.
  • Write design documents and drive cross-functional alignment on new features.
  • Engage with customers and the customer success team to understand performance requirements and use cases.

Requirements

  • 5+ years of software engineering experience.
  • Proficiency in C++.
  • Experience working with compilers for machine learning frameworks.
  • Creativity, curiosity, and the ability to collaborate effectively in a team-oriented environment.
  • Helpful qualifications include experience with MLIR, LLVM, ML parallel or distributed programming, heterogeneous ML computation, code generation, and ML frameworks such as PyTorch, JAX, or TensorFlow.

Benefits

  • Comprehensive healthcare coverage, retirement and savings programs, employee stock purchase opportunities, paid time off, wellbeing resources, family support programs, and learning and development opportunities may be included depending on location.
  • Competitive compensation packages may include RSU grants, annual target bonus, equity, and benefits.
  • Regular team onsites and local meetups are provided.
  • Candidates may work from the Los Altos, California office or remotely from home in the US or Canada.
  • New-hire onboarding is conducted in person at headquarters in Los Altos, California.
  • Traveling 2-4 times per year is expected for all roles.

Tech Stack

Categories

Modular

About Modular

201-500 employees

Modular builds an AI developer platform for training and especially inference/serving, centered on the MAX runtime and the Mojo programming language. Its tools accelerate and deploy models from frameworks like PyTorch and TensorFlow on CPUs and GPUs, for teams running on cloud or on‑prem infrastructure. Founded in 2022, the company operates remote‑first with an office in Los Altos, CA, and sells a commercial platform and enterprise support to organizations productionizing generative and classical ML.

Contact me