Anthropic

Anthropic

Red Team Engineer, Safeguards

Remote-Friendly (Travel Required) | San Francisco, CA

H1B sponsorship available$320k-$405kDetected 7 days ago
Distributed SystemsMachine LearningNLPLLMsCybersecurityPenetration TestingLogisticsTest AutomationResearchCommunicationAdaptabilityBurp Suite

About the role

  • Anthropic's Safeguards team is seeking a Red Team Engineer to help ensure the safety of our deployed AI systems and products.
  • In this role, you'll take an adversarial approach to uncover vulnerabilities across our product ecosystem before they can be exploited by malicious actors.
  • Your work will span from technical infrastructure vulnerabilities on our products to emergent risks from advanced AI capabilities.

Responsibilities

  • Conduct comprehensive adversarial testing across Anthropic's product surfaces, developing creative attack scenarios that combine multiple exploitation techniques
  • Research and implement novel testing approaches for emerging capabilities, including agent systems, tool use, and new interaction paradigms
  • Design and execute "full kill chain" attacks that emulate real-world threat actors attempting to achieve specific malicious objectives
  • Build and maintain systematic testing methodologies that evaluate every aspect of our systems
  • Develop automated testing frameworks to enable continuous assessment at scale
  • Collaborate with Product, Engineering, and Policy teams to translate findings into concrete improvements
  • Experience building custom automation, including LLM-specific testing frameworks

Requirements

  • Experience in penetration testing, red teaming, or application security
  • Experience in model jailbreaking and testing large-scale agentic workflows for non-obvious prompt injection vectors
  • Strong technical skills in web application security, including hands-on expertise with security testing tools (e.g., Burp Suite, Metasploit, custom scripting frameworks)
  • A track record of discovering novel attack vectors and chaining vulnerabilities in creative ways
  • Strong written and verbal communication skills, with the ability to explain technical concepts to varied audiences
  • Years of experience required will correlate with the internal job level requirements for the position
  • Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
  • Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position

Nice to have

  • Understanding of AI safety considerations beyond traditional security, including modern guardrails against jailbreaks
  • Experience testing API security and rate-limiting systems
  • Background in testing business logic vulnerabilities and authorization bypass techniques
  • Background in anti-fraud, trust & safety, or abuse prevention systems
  • Familiarity with distributed systems and infrastructure security
  • Familiarity with abuse detection mechanisms and the ability to engineer novel bypasses

Compensation

  • $320,000 - $405,000 USD

Benefits

  • Minimum education: Bachelor's degree or an equivalent combination of education, training, and/or experience

Visa & Work Authorization

  • However, we aren't able to successfully sponsor visas for every role and every candidate.
  • But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
  • We do sponsor visas!

This listing is sourced directly from Anthropic's careers page and normalized into a canonical job model.