320K–485K USD / year

Staff+ Software Engineer, Account Creation

AIHybrid — San Francisco, CA | New York City, NY | Seattle, WA
Published on 2026-09-24
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.

About this role

The role involves building safety and oversight mechanisms for AI systems, focusing on monitoring models and preventing misuse. Engineers will develop systems to detect unwanted behaviors and implement enforcement actions.

About the company

Anthropic builds reliable and interpretable AI systems, including the Claude family of models, for businesses and developers. The company focuses on AI safety research and practical AI assistants for enterprise and consumer use.

The team

Safeguards team

Stack

PythonTypeScript

What you'll do

  • Develop monitoring systems to detect unwanted behaviors
  • Build abuse detection mechanisms
  • Surface abuse patterns to research teams
  • Create multi-layered defenses for safety mechanisms

What we're looking for

  • Bachelor's degree in Computer Science or equivalent
  • Proficiency in Python and TypeScript
  • Ability to work across the stack
  • Strong communication skills

Nice to have

  • 8+ years of software engineering experience
  • Experience with abuse detection and mitigation
  • Experience building trust and safety mechanisms
  • Experience with prompt engineering and adversarial inputs

Benefits

  • Competitive compensation
  • Generous vacation and parental leave
  • Flexible working hours
  • Optional equity donation matching
View original job post