Jobs · empllo

H

AI Model Policy Trainer, Content Risk - Remote US

Handshake · United States · Posted today

remotejunior45-55 USD
Apply on empllo

About the role

📋 Description Evaluate user requests and AI model responses involving violence, weapons, threats, and dark Distinguish fictional, educational, historical, and defensive violence from requests that seek Assess whether a model's response gives meaningful real-world capability, regardless of how the Distinguish expressions of anger, frustration, or dark humor from credible threats or crisis Select the most defensible classification when a case is genuinely ambiguous, and write concise Write and refine adversarial or borderline prompts that probe where a model draws the line 🎯 Requirements You have spent serious time in violent fiction as a writer, game master, game designer You have real-world exposure to violence and its consequences through military, law enforcement You have worked with people in distress through crisis lines, counseling, threat assessment You use AI tools heavily and have opinions about where they refuse too much, help too much, or miss You notice when one word, contextual detail, or change in intent materially affects the answer You can hold a strong opinion without becoming attached to being right

Read the full posting on empllo

FAQ

Is the AI Model Policy Trainer, Content Risk - Remote US role at Handshake remote?+

This AI Model Policy Trainer, Content Risk - Remote US position is listed as remote (United States).

What is the salary for the AI Model Policy Trainer, Content Risk - Remote US role at Handshake?+

The listing states 45-55 USD.

What seniority level is this AI Model Policy Trainer, Content Risk - Remote US role?+

This is a junior level position.

How do I apply for the AI Model Policy Trainer, Content Risk - Remote US role at Handshake?+

Use the "Apply on empllo" button to open the original posting on empllo, where you can submit your application directly to Handshake.