Base Salary
$218k - $273k/yr
Responsibilities
- Design and implement end-to-end agent systems combining LLM reasoning, tool use, memory, and control logic
- Build scalable and reliable agent architectures deployable across customers with varying data, tools, and constraints
- Develop evaluation frameworks, datasets, environments, and metrics for production agent performance and business impact
- Translate enterprise requirements into robust agent designs with product managers, customers, data annotators, and engineering teams
- Productionize planning, multi-step reasoning, tool use, and multi-agent techniques into maintainable and observable systems
- Own deployment, monitoring, failure analysis, and continuous improvement of agent systems
- Contribute to technical direction, architecture, and best practices for general agent development, with increasing Staff-level leadership scope
Requirements
- 5+ years of experience building and deploying machine learning or AI systems for real-world production use cases
- Bachelor’s and/or master’s degree in Computer Science, Machine Learning, AI, or equivalent practical experience
- Deep understanding of modern LLMs, prompt-, context-, and system-level optimization, and agentic system design
- Proficiency in Python and production-quality, testable, maintainable software development
- Experience integrating models with external tools, APIs, databases, and services
- Ability to balance research-driven approaches with pragmatic product constraints
- Strong communication skills and comfort working in customer-facing or cross-functional environments
- Preferred experience with OpenAI APIs, commercial or open-source LLMs, agent frameworks, orchestration layers, and workflow systems
- Preferred familiarity with evaluation, monitoring, and observability for LLM-powered systems
- Preferred experience deploying machine learning systems in cloud environments and operating them at scale
- Preferred experience fine-tuning or adapting foundation models with SFT, RLVR, and LoRA
Benefits
- Comprehensive health, dental, and vision coverage
- Retirement benefits
- Learning and development stipend
- Generous paid time off
- Potential commuter stipend
- Full-time position located in San Francisco, New York, or Seattle
- 90-day waiting period before reconsideration for the same role
Tech Stack
Categories
About Scale AI
Scale’s mission is to develop reliable AI systems for the world’s most important decisions. We provide the high-quality data and full-stack technologies that power the world’s leading models, and help enterprises and governments build, deploy, and oversee AI applications that deliver real impact. The Scale Generative AI Platform allows customers to build, evaluate, and control advanced AI agents and applications that continuously improve. The Scale Data Engine provides the technology to collect, curate, and annotate high-quality datasets. Through our Scale Labs, we test models with rigorous benchmarks and novel research to ensure breakthroughs translate into systems people can trust. Scale powers the most advanced LLMs and generative models in the world through RLHF, data generation and model evaluation. We work with industry leaders like Meta, Cisco, DLA Piper, Mayo Clinic, Time Inc., the Government of Qatar, and U.S. government agencies including the Army and Air Force.