2 months ago
Responsibilities
- Implement and optimize current machine learning models for performance and accuracy across thousands of accelerators.
- Test and evaluate internal software releases, provide feedback to software engineering teams, make code fixes, and conduct code reviews.
- Benchmark models and machine learning techniques to identify performance bottlenecks and improve efficiency.
- Design, implement, and evaluate experiments involving novel AI methods.
- Collaborate with Research, Software, and Product teams to define, build, and test next-generation AI hardware.
- Engage with the AI community and stay current with developments in AI.
Requirements
- Bachelor’s, Master’s, PhD, or equivalent experience in Machine Learning, Computer Science, Mathematics, Data Science, or a related field.
- Proficiency with deep learning frameworks such as PyTorch or JAX.
- Strong Python or C++ software development skills.
- Expertise in deep learning across model training, optimization, and evaluation.
- Experience with distributed training or inference of machine learning models across 64 or more accelerators.
- Ability to design, execute, and report on machine learning experiments.
- Deep understanding of performance bottlenecks and methods to overcome them.
- Desirable experience with MLOps for Kubernetes-based clusters, production systems using large language models, or low-precision arithmetic.
- Desirable experience writing C++, Triton, or CUDA kernels for machine learning performance optimization.
- Familiarity with HPC systems and networking technologies including Infiniband, NVLink, and RoCE.
- Open-source contributions or published research papers are desirable.
- Knowledge of cloud computing platforms is desirable.
- Applicants must have the right to work in the UK; visa sponsorship is unavailable.
Benefits
- Flexible working arrangement
- Generous annual leave policy
- Private medical insurance and health cash plan
- Dental plan
- Pension matched up to 5%
- Life assurance and income protection
- Generous parental leave policy
- Employee assistance programme with health, mental wellbeing, and bereavement support
- Healthy food and snacks at the central Bristol office, including a barista bar
- Flexible interview approach and reasonable-adjustment support
- Applicants must have the right to work in the UK; visa sponsorship is not available
Tech Stack
Categories
About Graphcore
At Graphcore, we’re building the future of AI compute. We’re a team of semiconductor, software and AI experts, with deep experience in creating the complete AI compute stack - from silicon and software to infrastructure at datacenter scale. As part of the SoftBank Group, backed by significant long-term investment, we are delivering key technology into the fast-growing SoftBank AI ecosystem. To meet the vast and exciting AI opportunity, Graphcore is expanding its teams around the world. We are bringing together the brightest minds to solve the toughest problems, in a place where everyone has the opportunity to make an impact on the company, our products and the future of artificial intelligence.