about 3 hours ago
London, United KingdomSenior
Responsibilities
- Lead the design of infrastructure for moving GenAI ideas from prototype to production.
- Own and evolve the open-weights serving stack, including real-time GPU endpoints and batch inference.
- Architect scalable systems for model serving, batch inference, and fine-tuning.
- Push the cost and latency frontier of GPU inference, optimizing performance and reliability.
- Build platforms that support rapid experimentation while meeting production standards.
- Collaborate with ML engineers, product engineers, and data scientists to enhance platform capabilities.
- Set technical direction for the centralized GenAI platform, exploring new AI techniques.
Requirements
- BSc, MSc, or PhD in Computer Science or equivalent.
- 5+ years of industry experience in software engineering.
- Strong backend engineering fundamentals, especially in Python and distributed systems.
- Experience designing and owning production services, APIs, or ML infrastructure at scale.
- Hands-on experience with LLM inference and fine-tuning of open-weight models.
- Demonstrated technical leadership and mentoring experience.
- Proficiency in using AI coding tools throughout the software development lifecycle.