385K–460K USD / year

Product Manager, Safeguards (Account Integrity & Abuse)

ProductHybrid — San Francisco, CA | New York City, NY
Published on 2026-09-22
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.

About this role

As a Product Manager for the Safeguards team, you will be responsible for the ideation, design, development, and deployment of safety systems for AI products, ensuring they are safe and beneficial for users. The role involves collaboration with research and product teams to develop effective safeguards against misuse

About the company

Anthropic builds reliable and interpretable AI systems, including the Claude family of models, for businesses and developers. The company focuses on AI safety research and practical AI assistants for enterprise and consumer use.

The team

Safeguards team

What you'll do

  • Build in safety by design for AI products
  • Write safety evaluations and communicate on safety
  • Drive impact through prioritization and defining problems
  • Collaborate with cross-functional stakeholders
  • Plan for risk mitigation in AI deployment
  • Develop metrics to evaluate system performance

What we're looking for

  • Experience making technical tradeoff decisions
  • Strong understanding of user needs and safety concerns
  • Demonstrated ability in product strategy across teams
  • Experience in designing and building performance metrics
  • Strong decision-making skills in ambiguous situations
  • Ability to launch new products in a zero to one environment

Nice to have

  • 5+ years in product management
  • Experience in data, detection, and interventions
  • Ability to articulate complex concepts to non-technical audiences

Benefits

  • Competitive compensation
  • Generous vacation and parental leave
  • Flexible working hours
View original job post