about 7 hours ago
Responsibilities
- Design, build, and operate core components of a large-scale, real-time distributed decisioning system.
- Take end-to-end ownership of specific subsystems including architecture, implementation, testing, and deployment.
- Build a mental model of a highly distributed, complex production system and contribute effectively.
- Identify opportunities for continuous evolution of the system and tooling.
- Collaborate with engineers and applied scientists on distributed systems and ML-driven decision-making.
- Diagnose and resolve issues in a highly concurrent, multi-region distributed system.
- Mentor other engineers and promote engineering rigor and operational excellence.
- Communicate effectively with technical and non-technical stakeholders.
Requirements
- 10+ years of experience in designing and operating large-scale, low-latency distributed systems.
- Proven track record of owning complex systems end-to-end.
- Ability to quickly ramp up in unfamiliar, complex codebases.
- Experience in evolving and refactoring live, business-critical systems.
- Strong command of Java and knowledge of algorithms, data structures, and performance optimization.
- Practical experience with distributed caching and high-throughput messaging systems.
- Comfort operating in cloud infrastructure (GCP or AWS) at scale.
- Exposure to statistical or ML-driven techniques in production systems is a plus.
- Real-time bidding or ad-tech experience is a plus, as well as experience in adjacent high-stakes domains.
- B.S. or M.S. in Computer Science, Engineering, or equivalent experience.
- High ownership, self-motivation, and critical thinking skills.
- Collaborative and team-oriented mindset.
Benefits
- Global access to mental health and financial wellness support.
- Comprehensive healthcare benefits including medical, dental, and vision.
- Retirement options such as 401(k) or pension.
- Support for taking time off in accordance with local policies.