[Expression of Interest] Research Manager, Interpretability

Anthropic · San Francisco

Don’t apply blind. See how your CV matches Manager first — free, in 30 seconds.

You will leave NewLuxJob. We do not receive or handle applications.

Company
Anthropic
Location
San Francisco
Salary
$350,000 — $500,000
Posted
November 7, 2025

About this job

Note: we don't have open Research Manager positions on the Interpretability team at this time. However, we're actively growing our team of Research Engineers and Research Scientists. If you're excited about interpretability research and open to an individual contributor role, we encourage you to apply. About the Interpretability team When you see what modern language models are capable of, do you wonder, "How do these things work? How can we trust them?" The Interpretability team’s mission is to reverse engineer how trained models work, and Interpretability research is one of Anthropic’s core research bets on AI safety. We believe that a mechanistic understanding is the most robust way to make advanced systems safe. People mean many different things by "interpretability".

This is a short summary.

Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.

Most-requested skills in United States

Based on 636 United States vacancies that list requirements, these are the skills employers ask for most often.

  • Strategy57% of postings
  • Leadership41% of postings
  • Stakeholder Management30% of postings
  • Project Management23% of postings
  • Coaching11% of postings
  • Budgeting9% of postings
  • Agile/Scrum4% of postings
  • Change Management3% of postings
Manager skills

Similar jobs

See if your CV fits this job

Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.

  • Instant match score for this role
  • Tailored cover letter in one click
  • Free — no credit card
Check my CV — free
[Expression of Interest] Research Manager, Interpretability — Anthropic | NewLuxJob