mpathic2

mpathic2

Red Teaming Expert

Seattle, Washington, United States · Part-time

Sponsorship not specified$30k-$60kDetected 20 days ago
ExpressMachine LearningLLMsWriting

About the role

  • You will identify failure modes, edge cases, and policy gaps-particularly in scenarios involving distress, ambiguity, or escalation.
  • This role involves roleplaying and reviewing clinical scenarios with AI agents.
  • As such, we are ideally seeking candidates who bring creative or performance-driven strengths, as these competencies enhance the realism, nuance, and emotional depth needed for AI safety testing.

Responsibilities

  • We partner with leading technology companies to support red teaming, trust & safety, expert annotation, and model evaluation across high-stakes domains.
  • mpathic is seeking part-time, project-based Red Teaming Experts to support a red-teaming and evaluation campaign focused on AI safety and model behavior in sensitive, real-world interactions.
  • In this role, you will design, simulate, and evaluate conversations with AI systems to assess safety, risk, and behavioral performance.

Requirements

  • Professional experience in one or more of the following:
  • Strong understanding of:
  • Ability to identify:
  • Experience with or Interest in:
  • Evaluating AI-generated responses (no coding required, but must be tech-comfortable)
  • Nice to Have (Not Required)

Skills

  • Theatre degrees or studies
  • Acting, theatre, improv, or voice-over experience
  • Strong writing skills, especially dialogue or scenario writing
  • Experience creating or inhabiting characters (e.g., performers, TTRPG roleplay, narrative designers)
  • Conversational design, interaction writing, or scripted roleplay experience
  • Participation in gaming, interactive storytelling, or digital communities where roleplay is common
  • What You'll Be Working On
  • may include:
  • Designing and executing red-teaming scenarios across diverse user behaviors
  • Reviewing AI-generated responses for safety, accuracy, and policy compliance
  • Identifying failure modes, edge cases, and behavioral risks
  • Assessing whether AI appropriately recognizes and responds to distress or escalation

Compensation

  • $30-60/hour, depending on experience and specific project tasks/difficulty

Benefits

  • Mental health sensitivity, boundaries, and responsible AI behavior
  • Background in mental health, behavioral science, or psychology
  • Experience with AI systems in sensitive domains (e.g., healthcare, safety)

This listing is sourced directly from mpathic2's careers page and normalized into a canonical job model.