Skip to content
← Back to job listings

Red Team Engineer (Safeguards)

Anthropic · San Francisco, United States

External listingfull-time14 days ago

About The Role

Join Anthropic's Safeguards team as a Red Team Engineer, where you'll take an adversarial approach to uncover vulnerabilities in our AI systems and products. Your work will involve comprehensive adversarial testing, researching novel testing approaches, and collaborating with various teams to translate findings into concrete improvements. You'll also have access to a range of benefits, including comprehensive health insurance, paid parental leave, flexible paid time off, and more.

  • Conduct comprehensive adversarial testing across Anthropic's product surfaces, developing creative attack scenarios that combine multiple exploitation techniques.
  • Research and implement novel testing approaches for emerging capabilities, including agent systems, tool use, and new interaction paradigms.
  • Collaborate with Product, Engineering, and Policy teams to translate findings into concrete improvements.
  • A track record of discovering novel attack vectors and chaining vulnerabilities in creative ways
  • Experience building custom automation, including LLM-specific testing frameworks
  • Strong technical skills in web application security, including hands-on expertise with security testing tools (e.g., Burp Suite, Metasploit, custom scripting frameworks)
  • Strong written and verbal communication skills, with the ability to explain technical concepts to varied audiences
  • A public body of work such as CVEs, blog posts, or disclosed bug bounty reports
  • Experience in model jailbreaking and testing large-scale agentic workflows for non-obvious prompt injection vectors
  • Experience in penetration testing, red teaming, or application security
  • Familiarity with abuse detection mechanisms and the ability to engineer novel bypasses
  • Adaptability to understand and build engagements around emerging threats outside your direct area of expertise
  • Understanding of AI safety considerations beyond traditional security, including modern guardrails against jailbreaks
  • Experience with AI/ML security or adversarial machine learning
  • Familiarity with distributed systems and infrastructure security
  • Background in testing business logic vulnerabilities and authorization bypass techniques
  • We encourage you to apply even if you do not believe you meet every single qualification
  • Background in anti-fraud, trust & safety, or abuse prevention systems
  • Experience testing API security and rate-limiting systems

This is an external listing. JobSpring does not represent or verify the employer. Report this listing