Description
Handshake AI is hiring an AI Safety Policy Evaluator focused on Violence & Threats to evaluate user requests, AI model responses, and conversation history involving violence, weapons, threats, and dark fiction. The role distinguishes fictional, educational, historical, defensive, and defensive violence from requests seeking real-world capability or intent to harm, assesses borderline cases, writes rationales, probes policy boundaries, identifies gaps, and participates in calibration. It is an ongoing full-time W-2 assignment with TCWGlobal supporting Handshake AI, onsite in Seattle, Monday through Friday, with a planned start date of September 21, 2026, and compensation of $55-$120 per hour.
