3 hours ago
San Francisco, CA, USA or New York, NY, USAMid Level
H1B Sponsor
Base Salary
$1 - $2/yr
Responsibilities
- Design and execute reliable fine-tuning projects that deliver customized AI solutions for critical customers.
- Identify domains where Claude should improve and collaborate with Research teams on evaluations, reinforcement-learning environments, and training infrastructure.
- Optimize fine-tuning strategies, design evaluation frameworks, and contribute to novel training approaches and model-improvement methodologies.
- Partner with account executives and customers to translate requirements into immediate solutions and longer-term research opportunities.
- Serve as the primary technical advisor for customers on fine-tuning and model-improvement projects, including integration, deployment, and best practices.
- Collaborate with Sales, Product, Research, and Engineering teams and communicate complex technical concepts to technical and non-technical audiences.
- Travel occasionally to customer sites for workshops, research collaboration, and implementation support.
Requirements
- At least 3 years of experience training or fine-tuning deep-learning models.
- Experience in a customer- or client-facing role.
- Bachelor’s degree or equivalent combination of education, training, and experience; the field should be relevant through coursework, training, or professional experience.
- Advanced degree in Computer Science, Machine Learning, Artificial Intelligence, Statistics, or a related technical field is preferred.
- Experience designing evaluations, building reinforcement-learning environments, or contributing to model-training pipelines.
- Strong proficiency in at least one programming language, with Python preferred.
- Recent experience building production systems with large language models.
- Strong communication, interpersonal, collaboration, teaching, and mentoring skills, with the ability to work through ambiguity and competing priorities.
- Interest in developing safe and beneficial AI systems.
Benefits
- Flexible working hours, generous vacation, parental leave, and a collaborative office space.
- Optional equity donation matching.
- Current hybrid policy expects staff to work from an Anthropic office at least 25% of the time.
- Visa sponsorship may be available, with immigration-lawyer support.
- Occasional travel to customer sites is required.
Tech Stack
Categories
About Anthropic
We're an AI research company that builds reliable, interpretable, and steerable AI systems. Our first product is Claude, an AI assistant for tasks at any scale. Our research interests span multiple areas including natural language, human feedback, scaling laws, reinforcement learning, code generation, and interpretability.