Scale AI

AI Infrastructure Engineer, Sandbox Platform

Scale AI
Apply
2 days ago
Seattle, WA, USA +2 moreMid Level
H1B Sponsor

Base Salary

$180k - $225k/yr

Responsibilities

  • Design and build the sandboxing platform, client library, and API surface for secure code execution.
  • Ensure strong isolation, security, and reproducibility across user sessions and workloads.
  • Optimize cold-start latency, memory footprint, and resource utilization at scale.
  • Debug production issues, reduce error rates, monitor systems, and implement preventive fixes.
  • Partner with internal teams to understand platform needs, troubleshoot issues, and build supporting tooling.
  • Respond to incidents, conduct root-cause analysis, and address production reliability issues.
  • Help maintain the sandboxing product roadmap and balance immediate needs with long-term architecture.
  • Lead architecture reviews and own projects from design through deployment.

Requirements

  • At least 4 years of experience building high-performance systems software, including meaningful experience maintaining libraries, SDKs, or developer-facing APIs.
  • Deep understanding of Linux internals, including process isolation, memory management, cgroups, and namespaces.
  • Experience with containerization and virtualization technologies such as Docker, Firecracker, gVisor, QEMU, or Kata Containers.
  • Proficiency in a systems programming language such as Go, Rust, C, or C++.
  • Experience designing developer-facing APIs, propagating errors, writing documentation, and improving developer experience.
  • Comfort working across infrastructure layers from kernel modules to orchestration frameworks such as Kubernetes.
  • Strong debugging skills and the ability to navigate performance and security tradeoffs in production systems.
  • Ability to context-switch between incident response and proactive product development.
  • Preferred qualifications include infrastructure startup founder or early-engineer experience, familiarity with LLM agents and agent frameworks, secure multi-tenant workload experience, snapshotting and restore experience, open-source systems or developer-tools contributions, and production on-call experience.

Benefits

  • Base salary, equity, and benefits may be available for eligible roles.
  • Comprehensive health, dental, and vision coverage.
  • Retirement benefits, a learning and development stipend, and generous paid time off.
  • The role may be eligible for additional benefits such as a commuter stipend.
  • This is a full-time position located in San Francisco, New York, or Seattle, as specified by the posting subtitle.
  • The same role may not be reconsidered for 90 days after an application.
Scale AI

About Scale AI

501-1,000 employees

Scale’s mission is to develop reliable AI systems for the world’s most important decisions. We provide the high-quality data and full-stack technologies that power the world’s leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. The Scale Generative AI Platform allows customers to build, evaluate, and control advanced AI agents and applications that continuously improve. The Scale Data Engine provides the technology to collect, curate, and annotate high-quality datasets. Through our Scale Labs, we test models with rigorous benchmarks and novel research to ensure breakthroughs translate into systems people can trust. Scale powers the most advanced LLMs and generative models in the world through RLHF, data generation and model evaluation. We work with industry leaders like Meta, Cisco, DLA Piper, Mayo Clinic, Time Inc., the Government of Qatar, and U.S. government agencies including the Army and Air Force.

Contact me