mpathic2
Red Teaming Expert
Seattle, Washington, United States · Part-time
Sponsorship not specified$30k-$60kDetected 20 days ago
ExpressMachine LearningLLMsWriting
About the role
- You will identify failure modes, edge cases, and policy gaps-particularly in scenarios involving distress, ambiguity, or escalation.
- This role involves roleplaying and reviewing clinical scenarios with AI agents.
- As such, we are ideally seeking candidates who bring creative or performance-driven strengths, as these competencies enhance the realism, nuance, and emotional depth needed for AI safety testing.
Responsibilities
- We partner with leading technology companies to support red teaming, trust & safety, expert annotation, and model evaluation across high-stakes domains.
- mpathic is seeking part-time, project-based Red Teaming Experts to support a red-teaming and evaluation campaign focused on AI safety and model behavior in sensitive, real-world interactions.
- In this role, you will design, simulate, and evaluate conversations with AI systems to assess safety, risk, and behavioral performance.
Requirements
- Professional experience in one or more of the following:
- Strong understanding of:
- Ability to identify:
- Experience with or Interest in:
- Evaluating AI-generated responses (no coding required, but must be tech-comfortable)
- Nice to Have (Not Required)
Skills
- Theatre degrees or studies
- Acting, theatre, improv, or voice-over experience
- Strong writing skills, especially dialogue or scenario writing
- Experience creating or inhabiting characters (e.g., performers, TTRPG roleplay, narrative designers)
- Conversational design, interaction writing, or scripted roleplay experience
- Participation in gaming, interactive storytelling, or digital communities where roleplay is common
- What You'll Be Working On
- may include:
- Designing and executing red-teaming scenarios across diverse user behaviors
- Reviewing AI-generated responses for safety, accuracy, and policy compliance
- Identifying failure modes, edge cases, and behavioral risks
- Assessing whether AI appropriately recognizes and responds to distress or escalation
Compensation
- $30-60/hour, depending on experience and specific project tasks/difficulty
Benefits
- Mental health sensitivity, boundaries, and responsible AI behavior
- Background in mental health, behavioral science, or psychology
- Experience with AI systems in sensitive domains (e.g., healthcare, safety)
Apply directly at mpathic2 →Create a free account for alerts like thisView mpathic2 immigration profile
This listing is sourced directly from mpathic2's careers page and normalized into a canonical job model.