OpenAI

Software Engineer, Inference - Performance Optimization

OpenAI
Apply
4 months ago

Base Salary

$295k - $555k/yr

Responsibilities

  • Build and refine performance models that translate microbenchmark results into cost-to-serve estimates.
  • Analyze inference workloads end to end across applications, models, and fleet infrastructure.
  • Enhance tooling to identify bottlenecks affecting latency and throughput.
  • Partner with other teams to turn performance insights into concrete improvements and project the effects of future changes on inference.

Requirements

  • Deep expertise in performance profiling, benchmarking, analysis, and optimization.
  • Ability to reason from first principles about distributed systems, model inference, and hardware efficiency.
  • Comfort working across application behavior, kernels, accelerators, networking, and fleet scheduling.
  • Interest in collaborating with engineering and research teams to improve real production systems.

Categories

OpenAI

About OpenAI

1,001-5,000 employees

OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. AI is an extremely powerful tool that must be created with safety and human needs at its core. OpenAI is dedicated to putting that alignment of interests first — ahead of profit. To achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. Our investment in diversity, equity, and inclusion is ongoing, executed through a wide range of initiatives, and championed and supported by leadership. At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.