5 months ago
Remote, EMEASenior
Responsibilities
- Research and experiment on methods for specializing foundational models for agentic use cases.
- Build and maintain data and training pipelines.
- Design, analyze, and iterate on training, fine-tuning, and data-generation experiments.
- Keep current with research and the state of the art in LLMs, alignment, synthetic data generation, and code generation.
- Write high-quality pragmatic code and contribute to team planning and technical communication.
Requirements
- Experience with large language models and LLM post-training.
- Deep knowledge of Transformers and strong deep-learning fundamentals.
- Knowledge of distributed training and a strong machine-learning and engineering background.
- Research experience, including proposing and evaluating novel research ideas.
- Familiarity with multiple areas including LLM fine-tuning and alignment, synthetic data generation, continual learning, RLVR, and code generation.
- Programming experience, strong algorithmic skills, Linux familiarity, and Python with PyTorch or Jax.
- Familiarity with LLM capabilities and limitations, modern code agents, and rapidly iterating environments.
- Prior non-ML programming experience, especially outside Python, is a nice-to-have.
Benefits
- Fully remote work with flexible hours.
- 37 days per year of vacation and holidays.
- Health insurance allowance for the employee and dependents.
- 16 weeks of flexible, full-pay parental leave.
- Well-being, always-be-learning, and home-office allowances.
- Company-provided equipment.
- Frequent team get-togethers and a diverse, inclusive, people-first culture.
- Monthly three-day in-person collaboration in Paris and annual longer off-sites.
Categories
AI ResearchML Engineering
