Jobs · empllo
AI Model Policy Trainer, Content Risk - Remote US
Handshake · United States · Posted today
About the role
📋 Description Evaluate user requests and AI model responses involving violence, weapons, threats, and dark Distinguish fictional, educational, historical, and defensive violence from requests that seek Assess whether a model's response gives meaningful real-world capability, regardless of how the Distinguish expressions of anger, frustration, or dark humor from credible threats or crisis Select the most defensible classification when a case is genuinely ambiguous, and write concise Write and refine adversarial or borderline prompts that probe where a model draws the line 🎯 Requirements You have spent serious time in violent fiction as a writer, game master, game designer You have real-world exposure to violence and its consequences through military, law enforcement You have worked with people in distress through crisis lines, counseling, threat assessment You use AI tools heavily and have opinions about where they refuse too much, help too much, or miss You notice when one word, contextual detail, or change in intent materially affects the answer You can hold a strong opinion without becoming attached to being right
Read the full posting on empllo →
FAQ
Is the AI Model Policy Trainer, Content Risk - Remote US role at Handshake remote?+
This AI Model Policy Trainer, Content Risk - Remote US position is listed as remote (United States).
What is the salary for the AI Model Policy Trainer, Content Risk - Remote US role at Handshake?+
The listing states 45-55 USD.
What seniority level is this AI Model Policy Trainer, Content Risk - Remote US role?+
This is a junior level position.
How do I apply for the AI Model Policy Trainer, Content Risk - Remote US role at Handshake?+
Use the "Apply on empllo" button to open the original posting on empllo, where you can submit your application directly to Handshake.