d-Matrix

Software Engineer, Staff - SIMD Kernels

d-Matrix
Apply
8 months ago
Santa Clara, CA, USAStaff+
H1B sponsor

Base Salary

$190k - $300k/yr

Responsibilities

  • Develop, enhance, and maintain software kernels for ML operators on next-generation AI hardware.
  • Productize the software stack for the company’s AI compute engine.
  • Build SDK solutions that are intuitive for developers to use.
  • Analyze and improve software performance.
  • Map algorithms and AI-framework computational graphs onto modern hardware architectures.
  • Navigate hardware-software co-design trade-offs and deliver scalable, high-quality software.

Requirements

  • MS or PhD in computer engineering, math, physics, or a related degree.
  • 5+ years of industry experience.
  • Strong understanding of computer architecture, data structures, system software, and machine-learning fundamentals.
  • Proficiency in C/C++ and Python development in a Linux environment using standard development tools.
  • Experience implementing algorithms in C/C++ and Python.
  • Experience implementing algorithms for specialized hardware such as FPGAs, DSPs, GPUs, and AI accelerators using libraries such as CUDA.
  • Experience implementing ML operators including GEMMs, convolutions, softmax, layer normalization, and pooling.
  • Prior startup, small-team, or incubation experience is preferred.
  • Experience with TensorFlow, PyTorch, ML compilers such as MLIR, LLVM, TVM, or Glow, and deep-learning models for computer vision, NLP, or recommendation is preferred.
  • Experience developing for embedded SIMD vector processors such as Tensilica, or working at a cloud provider or AI compute/subsystem company, is preferred.

Benefits

  • Santa Clara, California headquarters or regional office location; remote work is possible.

Tech Stack

Categories

d-Matrix

About d-Matrix

201-500 employees

d-Matrix builds AI inference computing platforms for data centers, combining custom silicon with systems, networking, and software. Its flagship Corsair platform and JetStream fabric focus on low-latency, energy-efficient generative AI inference at scale. Founded in 2019 and headquartered in Santa Clara, California, the privately held company sells hardware with accompanying software to cloud providers and enterprises deploying large AI models.

Contact me