320K–485K USD / year
Staff+ Software Engineer, Account Creation
AIHybrid — San Francisco, CA | New York City, NY | Seattle, WA
Published on 2026-09-24
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.
About this role
The role involves building safety and oversight mechanisms for AI systems, focusing on monitoring models and preventing misuse. Engineers will develop systems to detect unwanted behaviors and implement enforcement actions.
About the company
Anthropic builds reliable and interpretable AI systems, including the Claude family of models, for businesses and developers. The company focuses on AI safety research and practical AI assistants for enterprise and consumer use.
The team
Safeguards team
Stack
PythonTypeScript
What you'll do
- Develop monitoring systems to detect unwanted behaviors
- Build abuse detection mechanisms
- Surface abuse patterns to research teams
- Create multi-layered defenses for safety mechanisms
What we're looking for
- Bachelor's degree in Computer Science or equivalent
- Proficiency in Python and TypeScript
- Ability to work across the stack
- Strong communication skills
Nice to have
- 8+ years of software engineering experience
- Experience with abuse detection and mitigation
- Experience building trust and safety mechanisms
- Experience with prompt engineering and adversarial inputs
Benefits
- Competitive compensation
- Generous vacation and parental leave
- Flexible working hours
- Optional equity donation matching
