Talent Apply
Log in
All jobs
E

ML evals engineer

Exa
onsiteUSD 180,000 - 350,000 / year
San Francisco, California

About this role

About the Role Join Exa's ML team to design and build the eval stack for next-generation search in an LLM-enabled world. You will create comprehensive evaluation methods to quantify search quality and inform research and product direction. Your work will directly influence how the company builds its core models and systems. What You'll Do

  • Write a manifesto of what perfect search means
  • Design and implement evaluation frameworks that probe the limits of search
  • Build scalable, reliable eval pipelines that track regressions, drift, and quality signals across billions of documents
  • Create golden datasets, synthetic benchmarks, agentic tasks, and real-world test suites that reflect how developers, agents, and humans actually use Exa
  • Partner closely with ML researchers, data engineers, infra engineers, and product to shape the feedback loops that improve our search models

Read the full description on TalentApply

Create a free account to see the complete job description, how well your CV matches this role, and apply in one click.

Clean up your CV

AI rewrites and formats your CV so it reads well and gets past screeners.

See how you score

Get your match percentage for this exact role before you spend time applying.

Apply professionally

Send a polished application in one click — no retyping the same details.

Track it easily

Follow every application in one place instead of digging through your inbox.

Free account · No card required

Your next opportunity starts here

Prepare, apply, track, interview and get hired — all from one platform, with AI in your corner.

Download app