AI Model Policy Trainer, Content Risk - Seattle Onsite

Added
14 minutes ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

content risk ai policy trainer

πŸ“‹ Description

  • Evaluate user requests and AI model responses involving violence, weapons, threats, and dark
  • Distinguish fictional, educational, historical, and defensive violence from requests that seek
  • Assess whether a model's response gives meaningful real-world capability, regardless of how the
  • Distinguish expressions of anger, frustration, or dark humor from credible threats or crisis
  • Select the most defensible classification when a case is genuinely ambiguous, and write concise
  • Write and refine adversarial or borderline prompts that probe where a model draws the line

🎯 Requirements

  • Strong instincts about violence in at least one of: fiction, real-world, or people in distress
  • Ability to hold a strong opinion without becoming attached to being right
  • Explain judgment calls clearly enough that another person can audit your reasoning
  • Can separate personal views from the standard a customer has asked to apply
  • Remain careful and consistent during repetitive work with difficult material
  • Communicate clearly and precisely in writing

🎁 Benefits

  • Compensation: $45-55/hr
  • W-2 employment classification
  • Structured evaluation frameworks and professional guidelines, with exposure limits, content
  • Opportunity to shape how every career evolves in the AI economy at global scale
  • Partner with world-class AI labs, Fortune 500 partners, and top educational institutions
  • Work with engineers, scientists, operators from Palantir, Meta, Scale AI, and former YC founders
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Data Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Data Jobs

See more Data jobs β†’