Head of Policy Design, Societal Harms

Anthropic · San Francisco

Don’t apply blind. See how your CV matches this role first — free, in 30 seconds.

You will leave NewLuxJob. We do not receive or handle applications.

Company
Anthropic
Location
San Francisco
Salary
$330,000 — $395,000
Posted
August 28, 2026

About this job

About the role Anthropic's Safeguards organization builds the policies, evaluations, and detection and enforcement systems that define and hold the limits on how Claude can be used. In this role, you'll lead our policy design team, managing the teams responsible for radicalization, child safety, user well-being, harmful manipulation, and election integrity, among other harm areas. The team is responsible for understanding and defining the risks that come with engaging with Claude, how those risks materialize in the real world, and the mitigations needed to prevent them. As the manager, you'll work with your team to draw the boundaries between what is and is not allowed, then partner with research, product, and engineering to build the right interventions.

This is a short summary.

Want to know if you're a fit? Check your CV against this role — free, in 30 seconds.

Similar jobs

See if your CV fits this job

Paste your CV for an instant match score against this role — and get a tailored cover letter in one click.

  • Instant match score for this role
  • Tailored cover letter in one click
  • Free — no credit card
Check my CV — free
Head of Policy Design, Societal Harms — Anthropic | NewLuxJob