ML evals engineer
About this role
About the Role Join Exa's ML team to design and build the eval stack for next-generation search in an LLM-enabled world. You will create comprehensive evaluation methods to quantify search quality and inform research and product direction. Your work will directly influence how the company builds its core models and systems. What You'll Do
- Write a manifesto of what perfect search means
- Design and implement evaluation frameworks that probe the limits of search
- Build scalable, reliable eval pipelines that track regressions, drift, and quality signals across billions of documents
- Create golden datasets, synthetic benchmarks, agentic tasks, and real-world test suites that reflect how developers, agents, and humans actually use Exa
- Partner closely with ML researchers, data engineers, infra engineers, and product to shape the feedback loops that improve our search models
Read the full description on TalentApply
Create a free account to see the complete job description, how well your CV matches this role, and apply in one click.
AI rewrites and formats your CV so it reads well and gets past screeners.
Get your match percentage for this exact role before you spend time applying.
Send a polished application in one click — no retyping the same details.
Follow every application in one place instead of digging through your inbox.
Free account · No card required
Your next opportunity starts here
Prepare, apply, track, interview and get hired — all from one platform, with AI in your corner.