Staff+ Software Engineer, Infrastructure, Interpretability
Anthropic · San Francisco
Don’t apply blind. See how your CV matches Software Developer first — free, in 30 seconds.
You will leave NewLuxJob. We do not receive or handle applications.
- Company
- Anthropic
- Location
- San Francisco
- Salary
- $405,000 — $485,000
- Posted
- August 12, 2026
About this job
How can we trust them?" The Interpretability team at Anthropic works to understand what's actually happening inside trained models - and applies our best techniques to keep frontier AI safe as it rapidly improves. Think of us as doing "neuroscience" of neural networks using "microscopes" we build - or reverse-engineering neural networks like binary programs. More resources to learn about our work: Our Research blog - covering advances including Monosemantic Features and Circuits An Intro to Interpretability from our research lead, Chris Olah The Urgency of Interpretability from CEO Dario Amodei Engineering Challenges Scaling Interpretability - directly relevant to this role 60 Minutes segment - see a demo of tooling our team built New Yorker article - what it's like to work on one of AI's…
This is a short summary.
Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.
Most-requested skills in United States
Based on 440 United States vacancies that list requirements, these are the skills employers ask for most often.
- Python40% of postings
- Java26% of postings
- AWS25% of postings
- Kubernetes23% of postings
- React22% of postings
- TypeScript21% of postings
- SQL15% of postings
- Agile/Scrum13% of postings
Similar jobs
See if your CV fits this job
Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.
- Instant match score for this role
- Tailored cover letter in one click
- Free — no credit card