Base Salary
$295k - $555k/yr
Responsibilities
- Bring new AI technologies into production alongside machine learning researchers, engineers, and product managers.
- Enable advanced research through engineering support.
- Introduce techniques, tools, and architecture to improve inference performance, latency, throughput, and efficiency.
- Build visibility tools for identifying bottlenecks and instability, then implement solutions for priority issues.
- Optimize code and Azure VM fleets for efficient GPU hardware utilization.
Requirements
- At least five years of professional software engineering experience.
- Understanding of modern machine learning architectures and inference performance optimization.
- Familiarity with PyTorch, NVIDIA GPUs, NCCL, CUDA, InfiniBand, MPI, and NVLink, or the ability to gain that familiarity quickly.
- Experience architecting, building, observing, and debugging production distributed systems.
- Experience rebuilding or substantially refactoring production systems as scale increases.
- Ability to own problems end-to-end, work independently, and identify high-priority problems.
- Collaborative, humble, and team-oriented attitude.
Categories
About OpenAI
OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. AI is an extremely powerful tool that must be created with safety and human needs at its core. OpenAI is dedicated to putting that alignment of interests first — ahead of profit. To achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. Our investment in diversity, equity, and inclusion is ongoing, executed through a wide range of initiatives, and championed and supported by leadership. At OpenAI, we believe artificial intelligence has the potential to help people solve immense global challenges, and we want the upside of AI to be widely shared. Join us in shaping the future of technology.