OpenAI

Software Engineer, Caching Infrastructure

OpenAI
Apply
1 year ago

Base Salary

$230k - $385k/yr

Responsibilities

  • Design, build, and operate a multi-tenant caching platform used across inference, identity, quota, and product experiences.
  • Define the long-term vision and roadmap for caching as a core infrastructure capability.
  • Scale the caching platform automatically with workload while minimizing tail latency and supporting diverse use cases.
  • Collaborate with networking, observability, database, infrastructure, and product teams to meet platform needs.
  • Balance performance, durability, reliability, and cost in caching infrastructure design.

Requirements

  • 5+ years of experience building and scaling distributed systems, with a focus on caching, load balancing, or storage systems.
  • Deep expertise with Redis, Memcached, or similar caching solutions, including clustering, durability configurations, client-side connection patterns, and performance tuning.
  • Production experience with Kubernetes, service meshes such as Envoy, and autoscaling systems.
  • Strong judgment around latency, reliability, throughput, and cost in platform design.
  • Ability to balance pragmatic engineering with long-term technical excellence in a fast-paced environment.

Tech Stack

AmbassadorKubernetesRedis

Categories

OpenAI

About OpenAI

10,000+ employees

OpenAI builds and deploys large-scale AI models and tools—including ChatGPT, GPT-4–class models, DALL·E, and Whisper—sold via APIs and enterprise subscriptions to developers and businesses. It monetizes through usage-based API pricing and ChatGPT Plus/Team/Enterprise, and also reaches customers via Microsoft’s Azure OpenAI Service. Founded in 2015 and headquartered in San Francisco, it operates as a private partnership.

Contact me