6 months ago
Base Salary
$380k - $555k/yr
Responsibilities
- Build and scale retrieval infrastructure across indexing, serving, and query execution.
- Develop low-latency, high-throughput systems for real-time model interaction.
- Partner with research to productionize embedding and retrieval techniques.
- Support dense, sparse, and hybrid retrieval pipelines.
- Own system performance, reliability, and observability at scale.
- Collaborate with Pretraining, Inference, and Product teams to integrate retrieval end to end.
- Contribute to model-system interfaces for agentic workflows.
Requirements
- Experience building and scaling distributed systems.
- Background in search, retrieval, or indexing systems.
- Familiarity with embedding-based or ML-powered systems.
- Experience with performance optimization and production reliability.
- Ability to work across machine learning and systems boundaries.
- Ability to apply first-principles thinking in ambiguous problem spaces.
Categories
BackendData Engineering
About OpenAI
OpenAI builds and deploys large-scale AI models and tools—including ChatGPT, GPT-4–class models, DALL·E, and Whisper—sold via APIs and enterprise subscriptions to developers and businesses. It monetizes through usage-based API pricing and ChatGPT Plus/Team/Enterprise, and also reaches customers via Microsoft’s Azure OpenAI Service. Founded in 2015 and headquartered in San Francisco, it operates as a private partnership.
