Machine Learning Research Scientist, Evaluations

Scale AI · San Francisco, CA; Seattle, WA; New York

Don’t apply blind. See how your CV matches this role first — free, in 30 seconds.

You will leave NewLuxJob. We do not receive or handle applications.

Company
Scale AI
Location
San Francisco, CA; Seattle, WA; New York
Salary
$180,600 — $225,750
Posted
August 26, 2026

About this job

Scale works with the industry's leading AI labs to provide high quality data and accelerate progress in GenAI research. We are looking for Research Scientists and Research Engineers with expertise in LLM post-training (SFT, RLHF, reward modeling) and evaluation. This role is on the evaluation pod within the GenAI Research Organization and will focus on building benchmarks and diagnosing model failure modes in both text and multimodal modalities. In this role, you will develop rigorous evaluations and diagnostic methods that reveal where frontier models fail and why. You will collaborate with researchers and engineers to define best practices in evaluation-driven AI development.

This is a short summary.

Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.

Similar jobs

See if your CV fits this job

Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.

  • Instant match score for this role
  • Tailored cover letter in one click
  • Free — no credit card
Check my CV — free
Machine Learning Research Scientist, Evaluations — Scale AI | NewLuxJob