22 hours ago
Base Salary
$235k - $294k/yr
Responsibilities
- Own the end-to-end architecture, implementation, delivery, and evaluation of agent capabilities.
- Define new technical patterns, propose architecture, and lead implementation for problems without established approaches.
- Deploy internally developed and community models into production for government customers.
- Maintain and improve production models and agents through retraining, hyperparameter tuning, and architectural updates.
- Build agent evaluation benchmarks, LLM judges, and verifiers to improve system performance.
- Develop scalable machine learning infrastructure that automates and optimizes ML services.
- Partner with product and research teams to scope high-impact initiatives and upcoming product lines.
- Work directly with government users and subject-matter experts and translate findings into technical direction.
- Mentor engineers, review technical work, and advise managers on feasibility.
- Communicate technical tradeoffs to non-technical stakeholders and advocate for ML techniques across engineering and product teams.
- Work within security, compliance, classified-environment, compute, and correctness constraints.
- Travel approximately 10% for customer interaction and team needs.
Requirements
- At least 5 years of experience building and deploying applied ML systems in production.
- Extensive production experience with generative AI, agentic AI, natural language processing, deep learning, deep reinforcement learning, or computer vision.
- Experience owning architectural decisions and defending technical tradeoffs.
- Experience shipping agentic systems with production traffic and rigorous evaluation.
- Strong knowledge of algorithms, data structures, and object-oriented programming.
- Strong Python programming skills and experience with PyTorch or TensorFlow.
- Experience mentoring or reviewing other engineers.
- Active security clearance is required.
Benefits
- Base salary, equity, and benefits are available for eligible roles.
- Benefits include comprehensive health, dental, and vision coverage, retirement benefits, a learning and development stipend, generous paid time off, and potentially a commuter stipend.
- This is a full-time position based in Washington, DC, with approximately 10% travel.
- The role may involve classified, air-gapped, or IL5+ environments.
Tech Stack
Categories
About Scale AI
Scale AI builds data annotation services and AI development tools for enterprises and government agencies, sold as a platform and managed services. Its products include the Scale Generative AI Platform for building and evaluating agents and the Data Engine for collecting, curating, and labeling training data, including RLHF and model evaluation. Founded in 2016 and headquartered in San Francisco, the company is privately held and works across domains from computer vision to LLM applications.
