Staff+ Software Engineer, Infrastructure, Interpretability

Anthropic · San Francisco

Don’t apply blind. See how your CV matches Software Developer first — free, in 30 seconds.

You will leave NewLuxJob. We do not receive or handle applications.

Company
Anthropic
Location
San Francisco
Salary
$405,000 — $485,000
Posted
August 12, 2026

About this job

How can we trust them?" The Interpretability team at Anthropic works to understand what's actually happening inside trained models - and applies our best techniques to keep frontier AI safe as it rapidly improves. Think of us as doing "neuroscience" of neural networks using "microscopes" we build - or reverse-engineering neural networks like binary programs. More resources to learn about our work: Our Research blog - covering advances including Monosemantic Features and Circuits An Intro to Interpretability from our research lead, Chris Olah The Urgency of Interpretability from CEO Dario Amodei Engineering Challenges Scaling Interpretability - directly relevant to this role 60 Minutes segment - see a demo of tooling our team built New Yorker article - what it's like to work on one of AI's

This is a short summary.

Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.

Most-requested skills in United States

Based on 440 United States vacancies that list requirements, these are the skills employers ask for most often.

  • Python40% of postings
  • Java26% of postings
  • AWS25% of postings
  • Kubernetes23% of postings
  • React22% of postings
  • TypeScript21% of postings
  • SQL15% of postings
  • Agile/Scrum13% of postings
Software Developer skills

Similar jobs

See if your CV fits this job

Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.

  • Instant match score for this role
  • Tailored cover letter in one click
  • Free — no credit card
Check my CV — free
Staff+ Software Engineer, Infrastructure, Interpretability — Anthropic | NewLuxJob