Staff+ Software Engineer (Safeguards Review Tooling)
Anthropic · San Francisco, United States
About The Role
Join Anthropic, a leading AI safety and research company. As a Staff+ Software Engineer on the Safeguards Review Tooling team, you will play a crucial role in ensuring the safe development and deployment of our models and products. You will build the systems that our safety investigators use to investigate potential harms and take enforcement actions. This is a foundational role where you will own the tools our safety investigators rely on and the platform underneath those tools. You will also drive how we scale review through automation and partner closely with various teams to ensure our enforcement systems are effective, accurate, and trustworthy.
- Concevoir et développer des outils d'investigation, de révision et d'application pour les surfaces de plateforme de première et de tierce partie.
- Développer la couche de plateforme d'API réutilisables, de stockage de données et de services backend qui permet de mettre en place rapidement et en toute sécurité de nouveaux flux de travail de révision.
- Automatiser le processus de révision, y compris permettre aux examinateurs d'utiliser Claude de manière efficace et construire vers des flux de travail de révision assistés et pilotés par Claude.
- A technical background in full-stack or platform engineering, with the ability to engage deeply in architecture and design discussions
- Excellent communication skills, including the ability to explain technical tradeoffs to non-technical stakeholders
- Experience shipping internal tools or platforms with demanding operational users, and a track record of improving their workflows measurably
- Care about the societal impacts of AI and want your work to make powerful systems safer
- Experience working cross-functionally with non-engineering partners such as operations, policy, or legal teams
- 8+ years of industry software engineering experience
- Experience building trust and safety, integrity, fraud, or abuse-prevention tooling, or other systems supporting human review at scale
- Experience designing systems under strict privacy, compliance, or data governance constraints, such as zero data retention environments
- Experience integrating LLMs or agentic systems into operational workflows, or building human-in-the-loop automation — including using agentic coding tools (e.g., Claude Code) as a core part of your own development workflow
- Experience building developer platforms or extensible tooling frameworks that other teams build on top of
- Experience supporting enforcement or moderation systems across multiple product surfaces, including enterprise or cloud platform contexts
- A product-minded approach to internal users: you work directly with the people using your tools, watch where they struggle, and fix it
This is an external listing. JobSpring does not represent or verify the employer. Report this listing
JobSpring