Research Scientist, Interpretability
Anthropic · San Francisco
Don’t apply blind. See how your CV matches this role first — free, in 30 seconds.
You will leave NewLuxJob. We do not receive or handle applications.
- Company
- Anthropic
- Location
- San Francisco
- Salary
- $350,000 — $850,000
- Posted
- November 7, 2025
About this job
We’re looking for researchers and engineers to join our efforts. People mean many different things by "interpretability". We're focused on mechanistic interpretability, which aims to discover how neural network parameters map to meaningful algorithms. Some useful analogies might be to think of us as trying to do "biology" or "neuroscience" of neural networks using “microscopes” we build, or as treating neural networks as binary computer programs we're trying to "reverse engineer". A few places to learn more about our work and team at a high level are this introduction to Interpretability from our research lead, Chris Olah; a discussion of our work on the Hard Fork podcast produced by the New York Times, and this blog post (and accompanying video) sharing more about some of the engineering…
This is a short summary.
Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.
Similar jobs
See if your CV fits this job
Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.
- Instant match score for this role
- Tailored cover letter in one click
- Free — no credit card