[Expression of Interest] Research Manager, Interpretability
Anthropic · San Francisco
Don’t apply blind. See how your CV matches Manager first — free, in 30 seconds.
You will leave NewLuxJob. We do not receive or handle applications.
- Company
- Anthropic
- Location
- San Francisco
- Salary
- $350,000 — $500,000
- Posted
- November 7, 2025
About this job
Note: we don't have open Research Manager positions on the Interpretability team at this time. However, we're actively growing our team of Research Engineers and Research Scientists. If you're excited about interpretability research and open to an individual contributor role, we encourage you to apply. About the Interpretability team When you see what modern language models are capable of, do you wonder, "How do these things work? How can we trust them?" The Interpretability team’s mission is to reverse engineer how trained models work, and Interpretability research is one of Anthropic’s core research bets on AI safety. We believe that a mechanistic understanding is the most robust way to make advanced systems safe. People mean many different things by "interpretability".…
This is a short summary.
Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.
Most-requested skills in United States
Based on 636 United States vacancies that list requirements, these are the skills employers ask for most often.
- Strategy57% of postings
- Leadership41% of postings
- Stakeholder Management30% of postings
- Project Management23% of postings
- Coaching11% of postings
- Budgeting9% of postings
- Agile/Scrum4% of postings
- Change Management3% of postings
Similar jobs
See if your CV fits this job
Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.
- Instant match score for this role
- Tailored cover letter in one click
- Free — no credit card