2 months ago
Base Salary
$182k - $242k/yr
Responsibilities
- Generate and investigate research ideas addressing obstacles to continuous learning in production.
- Work with the OpenPipe team to validate research directions across real customer tasks.
- Train, evaluate, and deploy LLM and machine-learning models using substantial GPU compute.
- Contribute across the stack, including CUDA kernels, distributed training systems, and high-performance LLM tracing dashboards.
- Help shape the roadmap and priorities of a small, autonomous engineering team.
Requirements
- Bachelor's, master's, or PhD in computer science, machine learning, robotics, or a related technical field.
- At least 4 years of machine-learning experience focused on model training, or a PhD with at least 2 years of relevant industry experience.
- Strong programming skills in Python and hands-on experience with PyTorch or JAX.
- Strong understanding of LLM post-training techniques, including supervised fine-tuning, reinforcement learning, and on-policy distillation.
- Experience developing, evaluating, and deploying machine-learning models in production environments.
- Preferred qualifications include publications, open-source contributions, research impact in LLM post-training or agent learning, distributed training and GPU acceleration experience, and experience leading complex projects or mentoring engineers.
Benefits
- Medical, dental, and vision insurance fully paid by CoreWeave for US-based employees.
- Company-paid life insurance, voluntary supplemental life insurance, and short- and long-term disability insurance.
- Flexible Spending Account, Health Savings Account, tuition reimbursement, and eligibility to participate in the Employee Stock Purchase Program.
- Mental wellness benefits through Spring Health and family-forming support through Carrot.
- Paid parental leave, flexible full-service childcare support through Kinside, and a 401(k) with employer match.
- Flexible PTO, catered lunch at office and data center locations, and a casual work environment.
Tech Stack
Categories
AI ResearchML Engineering
About CoreWeave
CoreWeave is the Essential Cloud for AI. CoreWeave is a cloud purpose-built for scaling, supporting, and accelerating GenAI. We’re a comprehensive platform and strategic partner designed to tackle today—and tomorrow’s—challenges of deploying AI at scale. We manage the complexities of AI growth to make supercomputing accessible and push the limits of what’s possible. Our teams create modern solutions to support modern technology. Get the premier choice for working with GenAI workloads.
