Research Engineer, Interpretability
Anthropic · San Francisco
Don’t apply blind. See how your CV matches Engineer first — free, in 30 seconds.
You will leave NewLuxJob. We do not receive or handle applications.
- Company
- Anthropic
- Location
- San Francisco
- Salary
- $315,000 — $560,000
- Posted
- November 7, 2025
About this job
Think of us as doing "neuroscience" of neural networks using "microscopes" we build - or reverse-engineering neural networks like binary programs. More resources to learn about our work: Our research blog - covering advances including Monosemantic Features and Circuits An Introduction to Interpretability from our research lead, Chris Olah The Urgency of Interpretability from CEO Dario Amodei Engineering Challenges Scaling Interpretability - directly relevant to this role 60 Minutes segment - Around 8:07, see a demo of tooling our team built New Yorker article - what it's like to work on one of AI's hardest open problems Even if you haven’t worked on interpretability before, the infrastructure expertise is similar to what's needed across the lifecycle of a production language model:…
This is a short summary.
Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.
Most-requested skills in United States
Based on 82 United States vacancies that list requirements, these are the skills employers ask for most often.
- Project Management49% of postings
- CAD26% of postings
- Quality Assurance9% of postings
- PLC7% of postings
- Lean7% of postings
- FEM/FEA7% of postings
- SolidWorks6% of postings
- MATLAB5% of postings
Similar jobs
See if your CV fits this job
Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.
- Instant match score for this role
- Tailored cover letter in one click
- Free — no credit card