mpathic2

mpathic2

Financial Risk & Safety Specialist (AI Systems)

Seattle, Washington, United States · Part-time

Sponsorship not specified$30k-$200kDetected 20 days ago
NLPAccountingCommunication

About the role

  • In this role, you will review AI-generated responses and multi-turn conversations to identify risks related to financial guidance, inappropriate agreement (e.g., sycophancy), overconfidence, and failure to appropriately communicate uncertainty or limitations.
  • Instead, it focuses on evaluation and red teaming-specifically, adversarial thinking and expert judgment applied to AI outputs in simulated scenarios.
  • You will help identify, prevent, and characterize risks that emerge when users engage AI systems in financial and general inquiry contexts.

Responsibilities

  • mpathic is seeking part-time Financial Experts to support a red-teaming and quality assurance (QA) campaign focused on evaluating AI system behavior in consumer-facing financial interactions.
  • We partner with leading technology companies to support red teaming, trust & safety, expert annotation, and model evaluation across high-stakes domains.

Requirements

  • Professional experience in one or more of the following:
  • Strong understanding of:
  • Ability to identify:
  • Strong written communication skills and ability to clearly explain reasoning
  • Successful candidates are thoughtful, detail-oriented, and able to apply financial expertise to assess risk, uncertainty, and appropriateness in conversational AI systems.
  • Experience with or interest in:
  • Experience with fintech, digital finance products, or robo-advisors
  • Prior experience with AI evaluation, annotation, or safety work

Nice to have

  • Red teaming, adversarial testing, or safety evaluation of AI systems
  • Evaluating how systems fail under realistic user behavior
  • Comfort working with AI tools and conversational outputs
  • Ability to work remotely using Slack and standard productivity tools
  • Comfort with ambiguity, iteration, and feedback-driven workflows
  • Willingness to sign NDAs and work with sensitive content
  • Availability ~10 hours per week for 8 weeks (starting in mid-April), with occasional scheduled meetings
  • Nice to Have (Not Required)

Skills

  • Collaborating with interdisciplinary teams on AI safety, policy, and evaluation frameworks

Compensation

  • $30-200/hour, depending on experience and specific project tasks/difficulty

Benefits

  • Our reviewers bring deep expertise in behavioral analysis, conversational design, mental health, and increasingly, financial and enterprise decision-making contexts.

This listing is sourced directly from mpathic2's careers page and normalized into a canonical job model.