12 hours ago
Zürich, SwitzerlandSenior
Responsibilities
- Design, train, and deploy production ML models for retrieval, reranking, and search relevance.
- Build and optimize embedding-based indexing and large-scale retrieval systems.
- Develop models for crawling, data selection, and content understanding.
- Define quality metrics and build evaluation pipelines for agent-native search.
- Work on high-throughput systems operating at very large scale.
- Collaborate with engineering teams to integrate ML models into production services.
- Analyze latency, quality, and cost trade-offs.
- Apply state-of-the-art techniques in search, retrieval, and LLM-integrated systems.
- Contribute to product and architectural decisions.
- Participate in coding interviews as part of the hiring process.
Requirements
- 5+ years of experience in software engineering or applied machine learning.
- Strong programming skills in Python, Go, or C++.
- Proven experience deploying ML models in production systems.
- Hands-on experience with retrieval, ranking, recommendation, or similar machine learning problems.
- Strong understanding of machine learning and modern deep learning techniques.
- Experience with large-scale data systems and high-throughput environments.
- Ability to design evaluation frameworks and define meaningful model metrics.
- Experience with search systems or large-scale information retrieval is preferred.
- Familiarity with embeddings, transformers, and modern NLP systems is preferred.
- Experience with LLM-powered or agent-based systems is preferred.
- Open-source contributions, technical publications, conference talks, or competitive ML experience such as Kaggle are preferred.
- Strong problem-solving, collaboration, and product-oriented skills.
Benefits
- Competitive compensation.
- Career growth and learning opportunities.
- Flexibility and ownership.
- Collaborative and innovative culture.
- Opportunity to work on impactful AI projects.
- International environment and talented teams.
Categories
About Nebius
Nebius builds a full-stack AI cloud offering GPU compute, storage, and tools for training and deploying ML models for startups, enterprises, and research labs. It sells consumption-based cloud infrastructure (IaaS/PaaS) and managed services tailored to generative AI workloads, including large-scale model training and inference. The company is headquartered in Amsterdam and operates as an independent provider.
