Anthropic

Program Manager, Safeguards Policy, Enforcement, and Threat Intelligence

San Francisco, CA·$285K–330K

Safeguards (Trust & Safety)

About Anthropic

Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.

About the role

Anthropic's Safeguards organization builds the policies, evaluations, and detection and enforcement systems that define and hold the limits on how Claude can be used. Within Safeguards, our Policy, Enforcement, and Threat Intelligence teams work as one org: Policy Design defines what is and isn't allowed across harm areas, Enforcement detects and acts on violations, and Threat Intelligence investigates sophisticated misuse and emerging threats.

That work runs across a lot of teams and moves fast. A single decision can touch model training, detection systems, product launches, legal review, and external communications. As the org's program manager, you'll keep all of it moving: tracking the portfolio across all three teams, running the processes that connect their work to each other and to the rest of the company, and making sure the right people are aligned before decisions land.

Your work will include the operating rhythm (planning, priorities, status, and decision tracking), launch and release readiness across policy, enforcement, and threat intelligence workstreams, and the processes that support how we share our work externally, including blog posts, threat reports, and other publications. You'll work closely with the

Read the rest on job-boards.greenhouse.io