13 hours ago
Remote, EMEASenior
Responsibilities
- Lead customer transitions from proof of concept to production and drive time-to-production and time-to-value.
- Ensure customer AI workloads are deployed, stable, scalable, performant, reliable, and cost-efficient.
- Understand customer architectures and use cases and proactively identify and resolve technical bottlenecks.
- Work directly with customer engineering and ML teams as a trusted technical partner.
- Act as the primary technical contact for production issues and coordinate incident resolution with internal teams.
- Provide feedback to Product and Infrastructure teams and identify opportunities to optimize and expand customer usage.
Requirements
- Practical knowledge of inference frameworks such as vLLM, TensorRT, or similar.
- Solid understanding of cloud or infrastructure systems, distributed systems or high-load applications, and AI/ML workloads including LLM inference.
- Ability to troubleshoot and reason about system performance.
- Experience working directly with technical customers such as engineers or ML teams.
- Ability to communicate complex technical topics clearly and effectively.
- Strong ownership, proactive execution, structured prioritization, and ability to manage multiple customers and priorities.
- Experience with GPU workloads, AI infrastructure, solutions engineering, SRE, or B2B technical support is a bonus.
Benefits
- Competitive compensation and career growth and learning opportunities.
- Flexible work arrangement with remote work available from Europe.
- Flexibility, ownership, collaborative culture, international environment, and opportunity to work on impactful AI projects.
Categories
Forward Deployed
About Nebius
Nebius builds a full-stack AI cloud offering GPU compute, storage, and tools for training and deploying ML models for startups, enterprises, and research labs. It sells consumption-based cloud infrastructure (IaaS/PaaS) and managed services tailored to generative AI workloads, including large-scale model training and inference. The company is headquartered in Amsterdam and operates as an independent provider.
