ML/Research Engineer, Safeguards

Anthropic · San Francisco, CA | New York City

Don’t apply blind. See how your CV matches Engineer first — free, in 30 seconds.

You will leave NewLuxJob. We do not receive or handle applications.

Company
Anthropic
Location
San Francisco, CA | New York City
Salary
$350,000 — $500,000
Posted
October 9, 2025

About this job

About the role We are looking for ML Engineers and Research Engineers to help detect and mitigate misuse of our AI systems. As a member of the Safeguards ML team, you will build systems that identify harmful use—from individual policy violations to sophisticated, coordinated attacks—and develop defenses that keep our products safe as capabilities advance. You will also work on systems that protect user wellbeing and ensure our models behave appropriately across a wide range of contexts. This work feeds directly into Anthropic's Responsible Scaling Policy commitments. Responsibilities Develop classifiers to detect misuse and anomalous behavior at scale.

This is a short summary.

Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.

Most-requested skills in United States

Based on 82 United States vacancies that list requirements, these are the skills employers ask for most often.

  • Project Management49% of postings
  • CAD26% of postings
  • Quality Assurance9% of postings
  • PLC7% of postings
  • Lean7% of postings
  • FEM/FEA7% of postings
  • SolidWorks6% of postings
  • MATLAB5% of postings
Engineer skills

Similar jobs

See if your CV fits this job

Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.

  • Instant match score for this role
  • Tailored cover letter in one click
  • Free — no credit card
Check my CV — free
ML/Research Engineer, Safeguards — Anthropic | NewLuxJob