about 18 hours ago
London, United KingdomMid Level / Senior
Responsibilities
- Build infrastructure to transition GenAI prototypes to production.
- Develop real-time GPU endpoints and high-throughput batch inference systems.
- Design scalable systems for model serving and fine-tuning.
- Optimize GPU inference costs and latency.
- Create platforms that support rapid experimentation and meet production standards.
- Collaborate with ML engineers and product teams to enhance platform capabilities.
- Shape the future of the centralized GenAI platform.
Requirements
- BSc, MSc, or PhD in Computer Science or equivalent.
- 3+ years of industry experience in software engineering.
- Strong backend engineering skills, particularly in Python and distributed systems.
- Experience building production services and ML infrastructure at scale.
- Hands-on experience with LLM inference and fine-tuning in production.
- Ability to navigate ambiguous technical areas and develop reusable platform capabilities.
- Proficiency in AI coding tools throughout the software development lifecycle.