H

On-site

AI Model Policy Trainer, Mental Health (Seattle)

Handshake

Overview

Review user conversations and AI responses involving mental health, emotional distress, crisis, unusual beliefs, paranoia, hallucinations, mania, and related experiences Apply detailed policy taxonomies and evaluation criteria to classify user requests and model behavior; evaluate individual messages and changes across longer conversations

Benefits, from the posting itself

Hover a row for the exact line it came from. Nothing inferred.

Medical, dental, and vision insurancereceipt ⟳
“Eligible employees can enroll in medical coverage, including prescription and mental health benefits, dental and vision insurance,”
Mental health benefitsreceipt ⟳
“including prescription and mental health benefits, dental and vision insurance,”
Pretax commuter benefitsreceipt ⟳
“healthcare and dependent care flexible spending accounts, and pretax commuter benefits.”
401(k) with employer matchreceipt ⟳
“A 401(k) retirement plan with an employer match is available to eligible participants.”
PTO accrues from first dayreceipt ⟳
“PTO accrues from the first day of the assignment at one hour per 40 hours worked, with no waiting period to use accrued PTO.”

Key Responsibilities

  • Review user conversations and AI responses involving mental health, emotional distress, crisis, unusual beliefs, paranoia, hallucinations, mania, and related experiences
  • Apply detailed policy taxonomies and evaluation criteria to classify user requests and model behavior; evaluate individual messages and changes across longer conversations
  • Distinguish between validating a person’s emotions and validating an unsupported or potentially harmful explanation; assess whether responses reinforce or escalate harmful beliefs
  • Evaluate actions taken by AI systems through tools, including searches, messages, scheduling, purchases, file creation, and publication
  • Write clear rationales that connect each evaluation decision to evidence from the conversation and policy