Anthropic
Red Team Engineer, Safeguards
Remote-Friendly (Travel Required) | San Francisco, CA
H1B sponsorship available$320k-$405kDetected 7 days ago
Distributed SystemsMachine LearningNLPLLMsCybersecurityPenetration TestingLogisticsTest AutomationResearchCommunicationAdaptabilityBurp Suite
About the role
- Anthropic's Safeguards team is seeking a Red Team Engineer to help ensure the safety of our deployed AI systems and products.
- In this role, you'll take an adversarial approach to uncover vulnerabilities across our product ecosystem before they can be exploited by malicious actors.
- Your work will span from technical infrastructure vulnerabilities on our products to emergent risks from advanced AI capabilities.
Responsibilities
- Conduct comprehensive adversarial testing across Anthropic's product surfaces, developing creative attack scenarios that combine multiple exploitation techniques
- Research and implement novel testing approaches for emerging capabilities, including agent systems, tool use, and new interaction paradigms
- Design and execute "full kill chain" attacks that emulate real-world threat actors attempting to achieve specific malicious objectives
- Build and maintain systematic testing methodologies that evaluate every aspect of our systems
- Develop automated testing frameworks to enable continuous assessment at scale
- Collaborate with Product, Engineering, and Policy teams to translate findings into concrete improvements
- Experience building custom automation, including LLM-specific testing frameworks
Requirements
- Experience in penetration testing, red teaming, or application security
- Experience in model jailbreaking and testing large-scale agentic workflows for non-obvious prompt injection vectors
- Strong technical skills in web application security, including hands-on expertise with security testing tools (e.g., Burp Suite, Metasploit, custom scripting frameworks)
- A track record of discovering novel attack vectors and chaining vulnerabilities in creative ways
- Strong written and verbal communication skills, with the ability to explain technical concepts to varied audiences
- Years of experience required will correlate with the internal job level requirements for the position
- Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience
- Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position
Nice to have
- Understanding of AI safety considerations beyond traditional security, including modern guardrails against jailbreaks
- Experience testing API security and rate-limiting systems
- Background in testing business logic vulnerabilities and authorization bypass techniques
- Background in anti-fraud, trust & safety, or abuse prevention systems
- Familiarity with distributed systems and infrastructure security
- Familiarity with abuse detection mechanisms and the ability to engineer novel bypasses
Compensation
- $320,000 - $405,000 USD
Benefits
- Minimum education: Bachelor's degree or an equivalent combination of education, training, and/or experience
Visa & Work Authorization
- However, we aren't able to successfully sponsor visas for every role and every candidate.
- But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this.
- We do sponsor visas!
Apply directly at Anthropic →Create a free account for alerts like thisView Anthropic immigration profile
This listing is sourced directly from Anthropic's careers page and normalized into a canonical job model.