OpenAI

Tech Lead, Model Serving Efficiency

OpenAI
Apply
over 1 year ago

Base Salary

$380k - $380k/yr

Responsibilities

  • Lead engineering efforts to improve model serving, inference performance, and system efficiency
  • Provide technical mentorship and oversight for a junior engineering team
  • Optimize kernels and data movement to improve system throughput and reliability
  • Partner with research and product teams to ensure models perform effectively at scale
  • Design, build, and improve critical serving infrastructure for Sora’s growth and reliability needs
  • Engage in model design to ensure trained models meet deployment requirements

Requirements

  • Deep expertise in model performance optimization, particularly at the inference layer
  • Strong background in kernel-level systems, data movement, and low-level performance tuning
  • Ability to mentor and guide junior engineers without formal management responsibilities
  • Ability to set technical direction and drive complex initiatives to completion
  • Interest in scaling high-performing AI systems for real-world multimodal workloads

Benefits

  • Hybrid work model requiring 3 days in the San Francisco office per week
  • Relocation assistance for new employees
OpenAI

About OpenAI

10,000+ employees

OpenAI builds and deploys large-scale AI models and tools—including ChatGPT, GPT-4–class models, DALL·E, and Whisper—sold via APIs and enterprise subscriptions to developers and businesses. It monetizes through usage-based API pricing and ChatGPT Plus/Team/Enterprise, and also reaches customers via Microsoft’s Azure OpenAI Service. Founded in 2015 and headquartered in San Francisco, it operates as a private partnership.

Contact me