1 year ago
Base Salary
$230k - $385k/yr
Responsibilities
- Design, build, and operate a multi-tenant caching platform used across inference, identity, quota, and product experiences.
- Define the long-term vision and roadmap for caching as a core infrastructure capability.
- Scale the caching platform automatically with workload while minimizing tail latency and supporting diverse use cases.
- Collaborate with networking, observability, database, infrastructure, and product teams to meet platform needs.
- Balance performance, durability, reliability, and cost in caching infrastructure design.
Requirements
- 5+ years of experience building and scaling distributed systems, with a focus on caching, load balancing, or storage systems.
- Deep expertise with Redis, Memcached, or similar caching solutions, including clustering, durability configurations, client-side connection patterns, and performance tuning.
- Production experience with Kubernetes, service meshes such as Envoy, and autoscaling systems.
- Strong judgment around latency, reliability, throughput, and cost in platform design.
- Ability to balance pragmatic engineering with long-term technical excellence in a fast-paced environment.
Tech Stack
About OpenAI
OpenAI builds and deploys large-scale AI models and tools—including ChatGPT, GPT-4–class models, DALL·E, and Whisper—sold via APIs and enterprise subscriptions to developers and businesses. It monetizes through usage-based API pricing and ChatGPT Plus/Team/Enterprise, and also reaches customers via Microsoft’s Azure OpenAI Service. Founded in 2015 and headquartered in San Francisco, it operates as a private partnership.
