×
Register Here to Apply for Jobs or Post Jobs. X
More jobs:

AI Model Policy Trainer, Content Risk

Job in Seattle, King County, Washington, 98127, USA
Listing for: Jobtailor
Full Time position
Listed on 2026-09-12
Job specializations:
  • IT/Tech
    AI Evaluation
Salary/Wage Range or Industry Benchmark: 65000 - 90000 USD Yearly USD 65000.00 90000.00 YEAR
Job Description & How to Apply Below
  • Evaluate user requests and AI model responses involving violence, weapons, threats, and dark fiction in full conversation context
  • Distinguish fictional, educational, historical, and defensive violence from requests seeking real-world uplift or expressing real intent to harm
  • Assess whether model responses provide meaningful real-world capability
  • Distinguish anger, frustration, and dark humor from credible threats or crisis indicators
  • Select defensible classifications for ambiguous cases and write concise policy-based rationales
  • Write and refine adversarial or borderline prompts
  • Identify policy gaps, contradictions, and emerging edge cases and raise them with project leads and policy teams
  • Participate in calibration discussions and update judgments based on stronger reasoning
  • Apply customer policy consistently without substituting personal beliefs
  • Maintain accuracy and attention to detail across repeated evaluations involving graphic material
Requirements
  • Strong judgment about violence in fiction, the real world, or among people in distress
  • Ability to distinguish fictional, educational, historical, and defensive violence from real-world uplift or intent to harm
  • Ability to assess meaningful real-world capability in AI model responses
  • Ability to distinguish anger, frustration, or dark humor from credible threats or crisis indicators
  • Ability to write concise rationales citing policy language and conversation details
  • Ability to write and refine adversarial or borderline prompts
  • Ability to identify policy gaps, contradictions, and emerging edge cases
  • Ability to participate in calibration discussions and apply customer policy consistently
  • Accuracy and attention to detail during repetitive, feedback-heavy evaluations
  • Clear and precise written communication
  • Ability to engage carefully, responsibly, and sustainably with graphic material
  • A degree, a clearance, and a technical background are not required
  • Must be authorized to work lawfully in the United States for Handshake
Core Competencies

Demonstrates strong judgment and analytical skills in evaluating AI model responses related to violence and threats, while maintaining accuracy and attention to detail. Capable of writing concise rationales and engaging with graphic material responsibly.

Highest-signal resume keywords
  • Strong Judgment About Violence
  • Ability To Distinguish Fictional And Real-World Violence
  • Ability To Write Concise Rationales
  • Ability To Identify Policy Gaps
  • Clear And Precise Written Communication
Soft Skills
  • Attention To Detail
  • Analytical Skills
  • Engagement With Graphic Material
Industry Keywords
  • AI Model Evaluation
  • Policy-Based Rationales
  • Calibration Discussions
  • Crisis Indicators
  • Adversarial Prompts
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary