Position:
Engineering Manager, Safeguards Interventions
Company:
Anthropic
Compensation:
Annual Salary: $405,000 - $485,000 USD
Location:
United States, San Francisco, CA
Employment type:
Not specified
Work Arrangement:
Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time.
Short Summary:
The Safeguards team is responsible for ensuring our models and products are developed and deployed safely. We're looking for an Engineering Manager to lead the Interventions team, which manages safety systems and interventions across various platforms.
Responsibilities:
- Hands-on lead and grow a team of engineers; own roadmap, OKRs, and execution.
- Drive cross-functional work with ML Infra, Research, Product, Policy, and Legal - and with cloud partners for 3P deployment.
- Set the bar for when an intervention is good enough to ship - backed by measurement.
- Own production reliability for intervention and compliance systems: incident response, postmortems, SLOs, and verification processes.
Requirement:
- Have managed engineering teams shipping production ML or safety-enforcement systems.
- Have run high-stakes, compliance-adjacent production systems.
- Care about measurement and have built evaluations that prove system efficacy.
- Can drive ambiguous, multi-stakeholder tradeoffs to a decision and own the outcome.
- Care deeply about AI safety.
Preferred qualifications:
- Experience in trust & safety, integrity, or abuse-prevention engineering at scale.
- Experience with compliance-driven systems and their legal/policy interfaces.
- Have shipped systems across multiple cloud providers.
Benefits:
- Competitive compensation and benefits.
- Optional equity donation matching.
- Generous vacation and parental leave.
- Flexible working hours.
- Collaborative office space.