SambaNova Systems

Principal Compiler Engineer - ML Systems

SambaNova Systems
Apply
3 months ago
Remote, United StatesStaff+
H1B Sponsor

Base Salary

$200k - $275k/yr

Responsibilities

  • Lead compiler engineering, including standard methodologies, enterprise product insertion, and process evolution.
  • Drive compiler infrastructure and optimization algorithms for high-performance machine-learning model execution.
  • Work across compiler-stack layers, development teams, domain experts, customers, and the broader enterprise.
  • Analyze PyTorch and machine-learning models and determine how to map operations onto underlying hardware.
  • Develop, integrate, and implement products.
  • Support proposals in areas aligned with core team competencies.

Requirements

  • Bachelor’s or master’s degree in computer science, computer engineering, or equivalent experience.
  • 5–10 years of industry experience.
  • Deep theoretical understanding of compiler fundamentals.
  • Experience building and deploying software products.
  • Experience with common compiler development practices and methodologies.
  • Experience with TensorFlow or PyTorch is a plus.
  • Experience with MLIR is preferred.
  • Familiarity with machine-learning models and frameworks, accelerated computing, and dataflow architectures is preferred.
  • Interest in high-performance systems engineering and performance debugging.

Benefits

  • Base salary plus equity and benefits for US-based full-time employment.
  • Employer covers 95% of employee medical insurance premiums and 77% of dependent premiums.
  • Health Savings Account with employer contribution.
  • Dental, vision, short- and long-term disability, basic life, voluntary life, AD&D, and Flexible Spending Account options.
  • Headspace subscription, Gympass+ membership, One Medical membership, counseling services, and Employee Assistance Program.

Tech Stack

PyTorchTensorFlow

Categories

SambaNova Systems

About SambaNova Systems

201-500 employees

Welcome to SambaNova: Revolutionizing AI Capacity At SambaNova, we're empowering developers, enterprises, governments, and data centers to unlock their full AI potential. Our full-stack infrastructure, from chips to models, enables lightning-fast performance, low power consumption, and high-efficiency computing. Our Mission To give every developer, enterprise, government and data center absolute sovereignty over their own data, models and AI infrastructure – to future-proof the AI workloads that will power and scale tomorrow. Our Technology We give our customers the optionality to experience SambaNova through the cloud or on-premise. Samba Cloud delivers the fastest inferences on the largest open source models like Llama 4 and DeepSeek. Developers can get started building in minutes with our OpenAI compatible APIs. All customers start on the developer tier and when they need more capacity can scale into our enterprise tier. SambaStack is our on-premise offering which includes the system, the platform, and foundation models. These components combine into a powerful technology stack that delivers unparalleled performance, ease of use, accuracy, data privacy, and the ability to power every use case across the world's largest organizations. SambaManaged is a modular and ready-to-deploy AI cloud designed to deliver unmatched efficiency for data centers and cloud service providers. This solution allows organizations to quickly deploy advanced AI inference services—without the need for costly infrastructure upgrades or specialized expertise—in as little as 90 days. At the heart of SambaNova innovation is the Reconfigurable Dataflow Unit (RDU). Purpose built for AI workloads, the RDU takes advantage of a dataflow architecture and a three-tiered memory design. The three tiers of memory enable the platform to run hundreds of models on a single node and to switch between them in microseconds. In 2023, SambaNova released its 4th generation RDU chip, the SN40L.

Contact me