4 months ago
Base Salary
$310k - $500k/yr
Responsibilities
- Scope, prototype, and run behavioral evaluations of AI systems in response to policy and oversight needs.
- Build evaluations for government and civil society partners, including harmful-manipulation evaluations for the EU AI Office.
- Design and run privileged-access evaluations and external oversight exercises with frontier AI labs.
- Adapt behavioral evaluation pipelines to the contexts of civil society organizations and domain experts.
- Translate ambiguous external stakeholder needs into concrete, technically credible deliverables.
Requirements
- Hands-on experience designing and running AI evaluations, particularly behavioral or interactive evaluations involving multi-turn, agentic, or red-teaming contexts.
- Strong engineering judgment and the ability to determine when work is ready to ship.
- Experience in customer-facing, consulting, or forward-deployed roles.
- Experience running evaluations at scale or in a production context.
- Ability to balance the needs of AI researchers, domain experts, and senior decision makers.
- Strong communication skills, low ego, and openness to feedback.
Benefits
- Benefits are provided.
- The role is based in San Francisco with an enthusiastic preference for working together in person.
- The organization is open to sponsoring international visas.
Categories
Forward Deployed
About Transluce
Transluce is an independent nonprofit research lab in San Francisco, founded in 2024, that builds open-source tools to analyze and evaluate advanced AI systems. It develops automated assessments of behaviors like honesty, misreporting, and evaluation awareness on open-weight models, then works with frontier AI labs and governments to apply vetted procedures. The lab focuses on public-interest AI transparency, reliability, and scalable oversight.
