Modular

Modular

Visit websiteLinkedIn201-500 employees

Modular builds an AI developer platform for training and especially inference/serving, centered on the MAX runtime and the Mojo programming language. Its tools accelerate and deploy models from frameworks like PyTorch and TensorFlow on CPUs and GPUs, for teams running on cloud or on‑prem infrastructure. Founded in 2022, the company operates remote‑first with an office in Los Altos, CA, and sells a commercial platform and enterprise support to organizations productionizing generative and classical ML.

Open Positions at Modular

5 open positions

Remote, United StatesSenior
$167k - $242k/yr

Build developer tools that help engineers create, debug, profile, and optimize AI inference models and kernels across heterogeneous hardware. You’ll shape agent-native workflows and improve the MAX developer experience for Modular’s AI serving platform.

1 month ago
Remote, United StatesSenior
$198k - $286k/yr

Inference Optimization Engineer building the platform, tooling, and software optimizations that deliver state-of-the-art LLM inference performance across kernels, engines, GPUs, ASICs, and cloud infrastructure. The role combines deep performance engineering with customer workload analysis and cross-functional technical leadership.

3 months ago
Remote, United StatesSenior
$198k - $286k/yr

Senior AI Graph Compiler Engineer responsible for advancing Modular’s ML graph compiler across heterogeneous hardware platforms. The role combines compiler optimization, ML systems performance, and collaboration across kernels, serving, Mojo, and customer teams.

4 months ago
Remote, United StatesSenior
$180k - $270k/yr

Lead the design and optimization of high-performance GPU and accelerator kernels that power large-scale AI inference. This senior role combines low-level systems expertise, hardware-aware optimization, and close collaboration with compiler and runtime teams.

8 months ago
Remote, United StatesSenior
$167k - $273k/yr

Build and scale a vertically integrated cloud inference platform for large language models, combining distributed systems, backend engineering, and machine learning infrastructure. This role focuses on high-performance, multi-node inference deployments for enterprises and developers.

8 months ago
Contact me