Exa

Research, Evals

Exa
Apply
10 months ago

Base Salary

$150k - $300k/yr

Responsibilities

  • Define a vision for perfect search and investigate evaluation approaches for search engines in an LLM world.
  • Design and implement evaluation frameworks that probe the limits of search.
  • Build scalable, reliable evaluation pipelines tracking regressions, drift, and quality signals across billions of documents.
  • Create golden datasets, synthetic benchmarks, agentic tasks, and real-world test suites for developers, agents, and humans.
  • Partner with ML researchers, data engineers, infrastructure engineers, and product teams to improve search-model feedback loops.

Requirements

  • Hands-on machine learning experience training, fine-tuning, or evaluating models, ideally involving embeddings or LLMs.
  • Strong engineering fundamentals and the ability to build reliable systems.
  • Experience with Python, Rust, distributed pipelines, and GPU or cluster jobs.
  • Enjoys building evaluation sets, inspecting edge cases, and designing creative measurement strategies.

Benefits

  • In-person role in San Francisco.
  • International candidate sponsorship is available, including STEM OPT, OPT, H1B, O1, and E3.
  • Premium medical, dental, and vision healthcare benefits.
  • Fertility benefits.
  • Monthly wellness stipend.
Exa

About Exa

51-200 employees

Exa is an applied AI research lab organizing human knowledge, starting with the web. Exa Search is the highest quality search API at every latency. It accepts natural language queries and produces token-efficient results with citations. Exa Agent is a high-compute web agent that synthesizes information from across the web to handle the most complex deep research, list-building, and enrichment workflows. These products power Cursor, Cognition, Hubspot, OpenRouter, Monday.com, and agents built by over 400,000 developers. We are a worldwide team of engineers and researchers. We're hiring: exa.ai/careers